Brand: NVIDIA | Category: GPUs
SKU: 900-21001-0030-000 | Part #: 900-21001-0030-000 | MPN: 900-21001-0030-000
Contact for Pricing — Request a Quote
PCIe Gen5 x16 connectivity and 700W thermal design power determine server chassis and PSU compatibility—critical specifications for any AI infrastructure team planning large-scale deployments. The NVIDIA H200 NVL 141GB HBM3e PCIe Datacenter GPU (part number 900-21001-0030-000) combines massive on-GPU memory with exceptional bandwidth to accelerate the most demanding workloads in modern datacenters. Built on the NVIDIA Hopper architecture with 4th Generation Tensor Cores, this passive-cooled accelerator delivers sparse tensor performance up to 3,958 TFLOPS in FP8 precision, enabling rapid iteration on large language models and generative AI applications.
The H200's 141 GB HBM3e memory and 4.8 TB/s bandwidth dramatically reduce data movement bottlenecks, while Third-Generation NVLink provides 900 GB/s bidirectional connectivity for multi-GPU scaling. NVIDIA's support for FP8, FP16, BF16, TF32, FP64, and INT8 precisions ensures flexibility across training, inference, and scientific computing tasks. ECC memory and Multi-Instance GPU (MIG) capability add enterprise-grade reliability and resource isolation. This PCIe Add-In Card runs on NVIDIA's datacenter driver stack under Linux, integrating seamlessly into existing GPU-accelerated clusters. The part number 900-21001-0030-000 is the definitive SKU for procurement.
For pricing, availability, and integration support, contact Omnixon Global via RFQ.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | 900-21001-0030-000 |
| Part Number | 900-21001-0030-000 |
| Condition | New |
| Manufacturer Part Number | 900-21001-0030-000 |
| GPU Architecture | NVIDIA Hopper (GH100) |
| Form Factor | PCIe Add-In Card |
| GPU Memory | 141 GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| Interconnect | PCIe Gen5 x16 |
| NVLink | Third-Generation NVLink, 900 GB/s bidirectional |
| Tensor Core Generation | 4th Generation |
| Supported Precisions | FP8, FP16, BF16, TF32, FP64, INT8 |
| FP8 Tensor Core Performance | 3,958 TFLOPS (sparse) |
| FP16 Tensor Core Performance | 1,979 TFLOPS (sparse) |
| BF16 Tensor Core Performance | 1,979 TFLOPS (sparse) |
| FP64 Tensor Core Performance | 67 TFLOPS |
| TDP (Thermal Design Power) | 700W |
| Cooling | Passive (requires datacenter airflow) |
| ECC Memory Support | Yes |
| Multi-Instance GPU (MIG) Support | Yes |
| PCIe Generation | PCIe 5.0 x16 |
| Operating System Support | Linux (NVIDIA datacenter driver stack) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.