Brand: NVIDIA | Category: GPUs
SKU: NVID-93523B000000000 | Part #: 935-23B00-0000-000 | MPN: 935-23B00-0000-000
Contact for Pricing — Request a Quote
The NVIDIA GB200 NVL72 Grace Blackwell Superchip is a rack-scale AI computing platform that integrates 36 Grace Blackwell Superchips—each pairing one NVIDIA Grace CPU with two NVIDIA Blackwell B200 GPUs—into a single NVLink-connected rack system. The result is 72 Blackwell GPUs and 36 Grace CPUs operating as a unified, coherent compute fabric interconnected via fifth-generation NVLink at 1.8 TB/s of all-to-all GPU bandwidth within the rack. The platform delivers up to 1.4 exaflops of AI compute at FP4 precision, establishing a new class of infrastructure for the most demanding large-scale AI training and inference workloads.
The GB200 NVL72 leverages the Blackwell GPU architecture, which introduces a second-generation Transformer Engine, native FP4 and FP6 precision support, and a new Reliability, Availability, and Serviceability (RAS) Engine designed for continuous operation at scale. The Grace CPUs provide high-bandwidth, low-latency coherent memory access between CPU and GPU via NVLink-C2C at 900 GB/s per Superchip, eliminating PCIe bottlenecks. Each B200 GPU in the system incorporates 192 GB of HBM3e memory with 8 TB/s of memory bandwidth per GPU, enabling the platform to hold and serve extremely large model parameter sets entirely in high-bandwidth memory without offloading.
Designed explicitly for AI factory deployments, the GB200 NVL72 targets organizations running frontier large language model training, multi-trillion-parameter model inference, and national-scale AI infrastructure. The rack integrates liquid cooling as the standard thermal solution to manage the platform's power envelope in hyperscale and enterprise data center environments. Its NVLink Switch topology allows every GPU in the rack to communicate directly with every other GPU at full bandwidth, making it the preferred foundation for distributed deep learning jobs that cannot tolerate the latency or bandwidth constraints of traditional scale-out networking alone.
| Manufacturer | NVIDIA |
| Manufacturer Part Number | 935-23B00-0000-000 |
| Platform Name | NVIDIA GB200 NVL72 |
| Form Factor | Rack-scale (full rack) |
| GPUs per Rack | 72 × NVIDIA Blackwell B200 GPUs |
| CPUs per Rack | 36 × NVIDIA Grace CPUs (ARM Neoverse V2-based) |
| GPU Architecture | NVIDIA Blackwell |
| GPU Memory per B200 | 192 GB HBM3e |
| Total GPU Memory (Rack) | 13.5 TB HBM3e |
| GPU Memory Bandwidth per B200 | 8 TB/s |
| Total GPU Memory Bandwidth (Rack) | 576 TB/s |
| AI Compute (FP4, Rack) | 1.4 exaflops |
| AI Compute (FP8, Rack) | 720 petaflops |
| CPU-GPU Interconnect | NVLink-C2C at 900 GB/s per Grace Blackwell Superchip (bidirectional) |
| Intra-Rack GPU Interconnect | NVIDIA NVLink 5 (fifth generation), 1.8 TB/s all-to-all aggregate bandwidth |
| NVLink Switch Generation | NVLink Switch (fifth generation) |
| External Fabric Support | NVIDIA InfiniBand and Ethernet (for multi-rack scale-out) |
| Thermal Solution | Liquid cooling (required) |
| Transformer Engine Generation | Second generation (native FP4, FP6, FP8, BF16, FP16, FP32, TF32) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | NVID-93523B000000000 |
| Part Number | 935-23B00-0000-000 |
| Condition | New |
| Manufacturer Part Number | 935-23B00-0000-000 |
| Platform Name | NVIDIA GB200 NVL72 |
| Form Factor | Rack-scale (full rack) |
| GPUs per Rack | 72 × NVIDIA Blackwell B200 GPUs |
| CPUs per Rack | 36 × NVIDIA Grace CPUs (ARM Neoverse V2-based) |
| GPU Architecture | NVIDIA Blackwell |
| GPU Memory per B200 | 192 GB HBM3e |
| Total GPU Memory (Rack) | 13.5 TB HBM3e |
| GPU Memory Bandwidth per B200 | 8 TB/s |
| Total GPU Memory Bandwidth (Rack) | 576 TB/s |
| AI Compute (FP4, Rack) | 1.4 exaflops |
| AI Compute (FP8, Rack) | 720 petaflops |
| CPU-GPU Interconnect | NVLink-C2C at 900 GB/s per Grace Blackwell Superchip (bidirectional) |
| Intra-Rack GPU Interconnect | NVIDIA NVLink 5 (fifth generation), 1.8 TB/s all-to-all aggregate bandwidth |
| NVLink Switch Generation | NVLink Switch (fifth generation) |
| External Fabric Support | NVIDIA InfiniBand and Ethernet (for multi-rack scale-out) |
| Thermal Solution | Liquid cooling (required) |
| Transformer Engine Generation | Second generation (native FP4, FP6, FP8, BF16, FP16, FP32, TF32) |