Brand: NVIDIA | Category: GPUs
SKU: GB300-NVL72 | Part #: GB300-NVL72 | MPN: GB300-NVL72
Contact for Pricing — Request a Quote
The NVIDIA GB300 NVL72 Grace Blackwell Ultra Rack System is a full-rack AI supercomputer built around the GB300 Grace Blackwell Ultra architecture, integrating 72 NVIDIA Blackwell Ultra GPUs and 36 Grace Blackwell Ultra Superchips into a single, cohesive rack-scale unit. Each GB300 Superchip combines an NVIDIA Grace CPU with a Blackwell Ultra GPU die interconnected via NVLink-C2C, delivering extreme memory bandwidth and unified CPU-GPU memory addressability that eliminates traditional PCIe bottlenecks. The system is interconnected via NVLink 5 across all 72 GPUs, creating a unified 72-GPU fabric with aggregate GPU memory that scales to support the largest frontier AI models in a single rack domain.
Designed for the most demanding AI training and inference workloads, the GB300 NVL72 delivers substantial improvements in FP4 and FP8 tensor core throughput relative to prior Hopper-generation systems, enabling enterprises to train and serve large language models and multimodal foundation models at scale with reduced total infrastructure footprint. The system ships as a complete rack-scale reference design incorporating liquid cooling infrastructure, NVLink Switch systems, high-bandwidth networking readiness, and power delivery components rated for datacenter integration, allowing operators to deploy immediately within compatible datacenter environments.
The GB300 NVL72 is positioned as an enterprise and hyperscale datacenter platform targeting AI research institutions, cloud service providers, sovereign AI initiatives, and large enterprises building proprietary foundation models or deploying high-throughput inference infrastructure. Its NVLink-based unified GPU fabric enables tensor-parallel and pipeline-parallel model execution across all 72 GPUs without leaving the rack, making it particularly well suited to workloads that exceed the memory capacity of any single GPU or node.
| Manufacturer | NVIDIA |
| Manufacturer Part Number | GB300-NVL72 |
| Product Line | Grace Blackwell Ultra |
| Architecture | NVIDIA Blackwell Ultra |
| System Form Factor | Full rack (rack-scale system) |
| GPU Count per Rack | 72 × NVIDIA Blackwell Ultra GPUs |
| Superchip Count per Rack | 36 × GB300 Grace Blackwell Ultra Superchips |
| CPU Architecture | NVIDIA Grace (Arm-based) — 1 Grace CPU per Superchip |
| CPU-GPU Interconnect | NVLink-C2C (chip-to-chip, within each Superchip) |
| GPU-to-GPU Interconnect | NVLink 5 (across all 72 GPUs via NVLink Switch) |
| NVLink Switch Generation | NVLink 5 |
| GPU Memory Type | HBM3e |
| Cooling | Liquid cooling |
| Tensor Core Precision Support | FP4, FP8, FP16, BF16, TF32, FP32, INT8 |
| Target Workloads | Generative AI training, large-scale LLM inference, HPC, scientific computing |
| Deployment Environment | Enterprise datacenter, hyperscale cloud, sovereign AI infrastructure |
| Networking Readiness | Compatible with high-bandwidth InfiniBand and Ethernet scale-out networking |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | GB300-NVL72 |
| Part Number | GB300-NVL72 |
| Condition | New |
| Manufacturer Part Number | GB300-NVL72 |
| Product Line | Grace Blackwell Ultra |
| Architecture | NVIDIA Blackwell Ultra |
| System Form Factor | Full rack (rack-scale system) |
| GPU Count per Rack | 72 × NVIDIA Blackwell Ultra GPUs |
| Superchip Count per Rack | 36 × GB300 Grace Blackwell Ultra Superchips |
| CPU Architecture | NVIDIA Grace (Arm-based) — 1 Grace CPU per Superchip |
| CPU-GPU Interconnect | NVLink-C2C (chip-to-chip, within each Superchip) |
| GPU-to-GPU Interconnect | NVLink 5 (across all 72 GPUs via NVLink Switch) |
| NVLink Switch Generation | NVLink 5 |
| GPU Memory Type | HBM3e |
| Cooling | Liquid cooling |
| Tensor Core Precision Support | FP4, FP8, FP16, BF16, TF32, FP32, INT8 |
| Target Workloads | Generative AI training, large-scale LLM inference, HPC, scientific computing |
| Deployment Environment | Enterprise datacenter, hyperscale cloud, sovereign AI infrastructure |
| Networking Readiness | Compatible with high-bandwidth InfiniBand and Ethernet scale-out networking |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.