Brand: NVIDIA | Category: GPUs
SKU: 900-21010-0060-000 | Part #: 900-21010-0060-000 | MPN: 900-21010-0060-000
Contact for Pricing — Request a Quote
The NVIDIA H200 PCIe 141GB GPU is built on the NVIDIA Hopper architecture and represents a significant advance in GPU memory capacity and bandwidth for enterprise AI and high-performance computing workloads. At its core, the H200 PCIe utilizes the same GH100 Tensor Core GPU as the H100, while introducing HBM3e memory technology that delivers 141GB of on-chip memory capacity — nearly double that of the H100 SXM5 — along with substantially increased memory bandwidth. This expanded memory footprint enables the inference and training of very large language models and multimodal AI models that previously required multi-GPU or multi-node configurations.
The H200 PCIe connects to host servers via a PCIe Gen5 x16 interface, making it compatible with a broad range of standard enterprise server platforms without requiring NVLink Switch System infrastructure. The GPU retains the fourth-generation Tensor Cores and Transformer Engine found in the H100, delivering high throughput for FP8, FP16, BF16, TF32, and FP64 precision workloads. NVLink support enables peer-to-peer GPU communication within a server node, and the card supports NVIDIA's full software ecosystem including CUDA, cuDNN, TensorRT, and the NVIDIA AI Enterprise software suite.
Designed for deployment in enterprise data centers, cloud service provider infrastructure, and on-premises AI compute clusters, the H200 PCIe 141GB addresses the growing memory requirements of generative AI inference, scientific simulation, and large-scale data analytics. Its PCIe form factor broadens accessibility across standard rack-mount server designs, while the combination of Hopper compute performance and HBM3e memory bandwidth makes it a capable platform for both training smaller frontier models and serving large deployed models with reduced latency and higher throughput per GPU.
| Manufacturer | NVIDIA |
| Manufacturer Part Number | 900-21010-0060-000 |
| GPU Architecture | NVIDIA Hopper (GH100) |
| GPU Memory | 141GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| Interface | PCIe Gen5 x16 |
| NVLink | Yes (NVLink 4.0, up to 600 GB/s bidirectional per GPU) |
| FP64 Tensor Core Performance | 34 TFLOPS |
| FP32 Performance | 67 TFLOPS |
| TF32 Tensor Core Performance | 989 TFLOPS (with sparsity: 1,979 TFLOPS) |
| FP16 / BF16 Tensor Core Performance | 1,979 TFLOPS (with sparsity: 3,958 TFLOPS) |
| FP8 Tensor Core Performance | 3,958 TFLOPS (with sparsity: 7,916 TFLOPS) |
| Transformer Engine | Yes (4th Generation) |
| Tensor Cores | 4th Generation |
| Thermal Design Power (TDP) | 350W |
| Form Factor | PCIe Full Height, Full Length (FHFL) dual-slot |
| ECC | Yes (HBM3e memory with ECC support) |
| Multi-Instance GPU (MIG) | Yes (up to 7 instances) |
| CUDA Compute Capability | 9.0 |
| Supported Precision Formats | FP64, FP32, TF32, BF16, FP16, FP8, INT8 |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | 900-21010-0060-000 |
| Part Number | 900-21010-0060-000 |
| Condition | New |
| Manufacturer Part Number | 900-21010-0060-000 |
| GPU Architecture | NVIDIA Hopper (GH100) |
| GPU Memory | 141GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| Interface | PCIe Gen5 x16 |
| NVLink | Yes (NVLink 4.0, up to 600 GB/s bidirectional per GPU) |
| FP64 Tensor Core Performance | 34 TFLOPS |
| FP32 Performance | 67 TFLOPS |
| TF32 Tensor Core Performance | 989 TFLOPS (with sparsity: 1,979 TFLOPS) |
| FP16 / BF16 Tensor Core Performance | 1,979 TFLOPS (with sparsity: 3,958 TFLOPS) |
| FP8 Tensor Core Performance | 3,958 TFLOPS (with sparsity: 7,916 TFLOPS) |
| Transformer Engine | Yes (4th Generation) |
| Tensor Cores | 4th Generation |
| Thermal Design Power (TDP) | 350W |
| Form Factor | PCIe Full Height, Full Length (FHFL) dual-slot |
| ECC | Yes (HBM3e memory with ECC support) |
| Multi-Instance GPU (MIG) | Yes (up to 7 instances) |
| CUDA Compute Capability | 9.0 |
| Supported Precision Formats | FP64, FP32, TF32, BF16, FP16, FP8, INT8 |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.