Brand: NVIDIA | Category: GPUs
SKU: 900-21010-0030-000 | Part #: 900-21010-0030-000 | MPN: 900-21010-0030-000
Contact for Pricing — Request a Quote
The NVIDIA H100 NVL 94GB HBM2e Dual-GPU is a high-capacity accelerator module built on the NVIDIA Hopper architecture, engineered specifically for large-scale generative AI inference, training, and high-performance computing workloads in enterprise datacenter environments. Featuring dual H100 GPUs on a single NVLink board, the platform delivers a combined 94GB of HBM2e memory, enabling organizations to run exceptionally large language models and multi-modal AI workloads entirely within GPU memory without model partitioning overhead.
The H100 NVL leverages the fourth-generation NVLink interconnect to bind the two GPUs into a unified, high-bandwidth memory domain, providing up to 900 GB/s of GPU-to-GPU bandwidth. The Hopper architecture introduces the Transformer Engine, which dynamically applies FP8 and FP16 mixed-precision computation to dramatically accelerate transformer-based model inference while sustaining accuracy. Fourth-generation Tensor Cores further extend throughput across FP8, FP16, BF16, TF32, and INT8 precisions, making the platform well-suited across the full AI development lifecycle from exploratory research to production serving.
The 900-21010-0030-000 SKU is designed for PCIe-based server integration, offering broad compatibility with standard enterprise rack infrastructure without requiring proprietary NVLink Switch fabric chassis. This positions the H100 NVL as a practical choice for organizations scaling AI infrastructure incrementally, with each dual-GPU board delivering the combined compute density of two discrete H100 SXM5-class processors in a single PCIe form factor, supported by NVIDIA's mature software ecosystem including CUDA, cuDNN, TensorRT, and the full NGC catalog of enterprise-ready AI frameworks and containers.
| Manufacturer | NVIDIA |
| Manufacturer Part Number | 900-21010-0030-000 |
| Product Line | NVIDIA H100 NVL |
| Architecture | NVIDIA Hopper (GH100) |
| GPU Configuration | Dual-GPU (2× H100 per board) |
| Total GPU Memory | 94 GB HBM2e |
| Memory per GPU | 47 GB HBM2e |
| Memory Bandwidth (Total) | Up to 7.8 TB/s aggregate |
| NVLink Interconnect | 4th Generation NVLink, up to 900 GB/s GPU-to-GPU bandwidth |
| FP8 Tensor Core Performance | Up to 3,958 TFLOPS (sparse) per board |
| FP16 Tensor Core Performance | Up to 1,979 TFLOPS (sparse) per board |
| TF32 Tensor Core Performance | Up to 990 TFLOPS (sparse) per board |
| Form Factor | PCIe (dual-slot) |
| Host Interface | PCIe Gen 5 x16 |
| Transformer Engine | Yes (FP8 and FP16 mixed precision) |
| Multi-Instance GPU (MIG) | Supported |
| Confidential Computing | Supported |
| ECC Memory Support | Yes |
| Total Board Power (TDP) | 350 W |
| Operating Temperature | 0°C to 35°C (inlet) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | 900-21010-0030-000 |
| Part Number | 900-21010-0030-000 |
| Condition | New |
| Manufacturer Part Number | 900-21010-0030-000 |
| Product Line | NVIDIA H100 NVL |
| Architecture | NVIDIA Hopper (GH100) |
| GPU Configuration | Dual-GPU (2× H100 per board) |
| Total GPU Memory | 94 GB HBM2e |
| Memory per GPU | 47 GB HBM2e |
| Memory Bandwidth (Total) | Up to 7.8 TB/s aggregate |
| NVLink Interconnect | 4th Generation NVLink, up to 900 GB/s GPU-to-GPU bandwidth |
| FP8 Tensor Core Performance | Up to 3,958 TFLOPS (sparse) per board |
| FP16 Tensor Core Performance | Up to 1,979 TFLOPS (sparse) per board |
| TF32 Tensor Core Performance | Up to 990 TFLOPS (sparse) per board |
| Form Factor | PCIe (dual-slot) |
| Host Interface | PCIe Gen 5 x16 |
| Transformer Engine | Yes (FP8 and FP16 mixed precision) |
| Multi-Instance GPU (MIG) | Supported |
| Confidential Computing | Supported |
| ECC Memory Support | Yes |
| Total Board Power (TDP) | 350 W |
| Operating Temperature | 0°C to 35°C (inlet) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.