Brand: NVIDIA | Category: GPUs
SKU: 699-21010-0202-000 | Part #: 699-21010-0202-000 | MPN: 699-21010-0202-000
Contact for Pricing — Request a Quote
The NVIDIA H100 NVL is a dual-GPU PCIe module featuring two H100 GPUs with 94GB of HBM2e memory per GPU (188GB total), designed for high-performance compute in standard enterprise server chassis without requiring SXM5 sockets. Each GPU delivers 141.1 TFLOPS of FP8 throughput and 70.6 TFLOPS of FP32, connected via NVLink for peer-to-peer communication and optimized for distributed inference workloads.
Purpose-built for large language model inferencing at scale, the H100 NVL addresses supply constraints affecting SXM variants while maintaining architectural feature parity including Transformer Engine support, fourth-generation NVLink, and PCIe 5.0 connectivity. The dual-GPU module fits into standard PCIe slots, enabling retrofit deployment in existing data center infrastructure without server redesign.
Target deployments span telecommunications, financial services, and public sector organizations requiring high-throughput text generation, reasoning tasks, and multi-concurrent model serving. The 188GB aggregate memory footprint supports inference of large foundation models with extended context windows and batch processing optimization.
| Manufacturer | NVIDIA |
| ManufacturerPartNumber | 699-21010-0202-000 |
| ModelName | H100 NVL |
| FormFactor | Dual-GPU PCIe Module |
| GPUsPerModule | 2 |
| MemoryPerGPU | 94GB HBM2e |
| TotalMemory | 188GB |
| MemoryBandwidth | 3.35TB/s per GPU |
| PCIeInterface | PCIe 5.0 |
| NVLinkConnectivity | Fourth-generation NVLink (900GB/s per GPU) |
| FP32Performance | 70.6 TFLOPS per GPU |
| TensorFloat32Performance | 141.1 TFLOPS per GPU |
| FP8Performance | 565 TFLOPS per GPU |
| TransformerEngine | Supported |
| CudaComputeCapability | 9.0 |
| MaxPower | 700W per GPU |
| LaunchQuarter | 2023 Q4 |
| TargetApplication | LLM Inference |
| ServerCompatibility | Standard PCIe x16 capable servers |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | 699-21010-0202-000 |
| Part Number | 699-21010-0202-000 |
| Condition | New |
| ManufacturerPartNumber | 699-21010-0202-000 |
| ModelName | H100 NVL |
| FormFactor | Dual-GPU PCIe Module |
| GPUsPerModule | 2 |
| MemoryPerGPU | 94GB HBM2e |
| TotalMemory | 188GB |
| MemoryBandwidth | 3.35TB/s per GPU |
| PCIeInterface | PCIe 5.0 |
| NVLinkConnectivity | Fourth-generation NVLink (900GB/s per GPU) |
| FP32Performance | 70.6 TFLOPS per GPU |
| TensorFloat32Performance | 141.1 TFLOPS per GPU |
| FP8Performance | 565 TFLOPS per GPU |
| TransformerEngine | Supported |
| CudaComputeCapability | 9.0 |
| MaxPower | 700W per GPU |
| LaunchQuarter | 2023 Q4 |
| TargetApplication | LLM Inference |
| ServerCompatibility | Standard PCIe x16 capable servers |