Brand: NVIDIA | Category: GPUs
SKU: NVH200NVLTCGPU-KIT | Part #: NVH200NVLTCGPU-KIT | MPN: NVH200NVLTCGPU-KIT
Contact for Pricing — Request a Quote
The PNY NVIDIA H200 NVL is a PCIe-based GPU accelerator purpose-built for large-language model inference and generative AI workloads. It features 141GB of HBM3e memory with 4.8TB/s memory bandwidth, enabling efficient processing of massive transformer models with minimal latency. The NVL variant is optimized for inference-scale deployments, delivering substantial throughput improvements over prior-generation accelerators in production environments.
Architecturally, the H200 NVL integrates NVIDIA's Hopper generation GPU cores with cutting-edge HBM3e technology, providing the memory capacity and bandwidth required to serve large context windows and batch inference requests without model quantization. Support for Transformer Engine and tensor operations enables efficient FP8, FP16, and TF32 compute across diverse model architectures, making it suitable for real-time and near-real-time inference at enterprise scale.
| Manufacturer | NVIDIA |
| Brand | PNY |
| Model | H200 NVL |
| Manufacturer Part Number | NVH200NVLTCGPU-KIT |
| GPU Architecture | Hopper |
| Memory Capacity | 141GB |
| Memory Type | HBM3e |
| Memory Bandwidth | 4.8TB/s |
| Interface | PCIe 5.0 |
| Max Power Consumption | 700W |
| Tensor Cores | 14400 |
| Transformer Engine | Yes |
| NVLINK Support | No |
| Form Factor | PCIe Single-Slot |
| CUDA Capability | 9.0 |
| Supported Precision | FP8, FP16, TF32, BF16 |
| Launch Date | Q4 2024 |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | NVH200NVLTCGPU-KIT |
| Part Number | NVH200NVLTCGPU-KIT |
| Condition | New |
| Model | H200 NVL |
| Manufacturer Part Number | NVH200NVLTCGPU-KIT |
| GPU Architecture | Hopper |
| Memory Capacity | 141GB |
| Memory Type | HBM3e |
| Memory Bandwidth | 4.8TB/s |
| Interface | PCIe 5.0 |
| Max Power Consumption | 700W |
| Tensor Cores | 14400 |
| Transformer Engine | Yes |
| NVLINK Support | No |
| Form Factor | PCIe Single-Slot |
| CUDA Capability | 9.0 |
| Supported Precision | FP8, FP16, TF32, BF16 |
| Launch Date | Q4 2024 |