Brand: NVIDIA | Category: GPUs
SKU: 900-21010-0120-000 | Part #: 900-21010-0120-000 | MPN: 900-21010-0120-000
Contact for Pricing — Request a Quote
The NVIDIA H100 NVL 188GB Dual-GPU is a specialized accelerator pairing designed for large-scale language model inference and generative AI workloads. Each H100 NVL GPU features 141 streaming multiprocessors with 18,176 CUDA cores, 568 fourth-generation Tensor cores per SM, and 94GB of HBM3 memory per chip, delivering up to 60 teraflops of FP8 tensor performance per GPU. The dual-GPU configuration totals 188GB of memory and is optimized for inference at massive scale, leveraging Hopper architecture enhancements including Transformer Engine acceleration and Dynamic Tensor Memory optimization.
Designed for density-constrained environments where power efficiency and memory bandwidth are critical, the H100 NVL configuration enables sub-100W per-GPU thermal profiles while maintaining peak throughput for batched inference scenarios. The architecture supports 141GB/s aggregate memory bandwidth per GPU, enabling rapid token generation and multi-user inference serving across large parameter models. This configuration became available in Q2 2023 and has become foundational infrastructure for enterprise LLM deployment stacks.
The dual-GPU pairing is engineered for inference-optimized use cases where throughput per watt and memory-to-compute ratios directly impact deployment economics and model serving capacity.
| Manufacturer | NVIDIA |
| Model | H100 NVL |
| ManufacturerPartNumber | 900-21010-0120-000 |
| GPUCount | 2 |
| TotalMemory | 188GB |
| MemoryPerGPU | 94GB HBM3 |
| MemoryBandwidth | 282GB/s (141GB/s per GPU aggregate) |
| CUDACores | 36352 |
| TensorCoresPerSM | 568 |
| StreamingMultiprocessors | 282 |
| Architecture | NVIDIA Hopper |
| MaxPower | 250W |
| NVLinkBandwidth | 900GB/s (per connection) |
| FP8TensorPerformance | 120 teraflops (aggregate dual-GPU) |
| TF32TensorPerformance | 60 teraflops (aggregate dual-GPU) |
| TransformerEngineSupport | true |
| LaunchQuarter | Q2 2023 |
| InterconnectSupport | NVLink 4.0, PCIe Gen5 |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | 900-21010-0120-000 |
| Part Number | 900-21010-0120-000 |
| Condition | New |
| Model | H100 NVL |
| ManufacturerPartNumber | 900-21010-0120-000 |
| GPUCount | 2 |
| TotalMemory | 188GB |
| MemoryPerGPU | 94GB HBM3 |
| MemoryBandwidth | 282GB/s (141GB/s per GPU aggregate) |
| CUDACores | 36352 |
| TensorCoresPerSM | 568 |
| StreamingMultiprocessors | 282 |
| Architecture | NVIDIA Hopper |
| MaxPower | 250W |
| NVLinkBandwidth | 900GB/s (per connection) |
| FP8TensorPerformance | 120 teraflops (aggregate dual-GPU) |
| TF32TensorPerformance | 60 teraflops (aggregate dual-GPU) |
| TransformerEngineSupport | true |
| LaunchQuarter | Q2 2023 |
| InterconnectSupport | NVLink 4.0, PCIe Gen5 |