Brand: NVIDIA | Category: GPUs
SKU: NVID-93523B000000000 | Part #: 935-23B00-0000-000 | MPN: 935-23B00-0000-000
Contact for Pricing — Request a Quote
The NVIDIA GB200 NVL72 Grace Blackwell Superchip is a rack-scale AI computing platform that integrates 36 Grace Blackwell Superchips—each pairing one NVIDIA Grace CPU with two NVIDIA Blackwell B200 GPUs—into a single NVLink-connected rack system. The result is 72 Blackwell GPUs and 36 Grace CPUs operating as a unified, coherent compute fabric interconnected via fifth-generation NVLink at 1.8 TB/s of all-to-all GPU bandwidth within the rack. The platform delivers up to 1.4 exaflops of AI compute at FP4 precision, establishing a new class of infrastructure for the most demanding large-scale AI training and inference workloads.
The GB200 NVL72 leverages the Blackwell GPU architecture, which introduces a second-generation Transformer Engine, native FP4 and FP6 precision support, and a new Reliability, Availability, and Serviceability (RAS) Engine designed for continuous operation at scale. The Grace CPUs provide high-bandwidth, low-latency coherent memory access between CPU and GPU via NVLink-C2C at 900 GB/s per Superchip, eliminating PCIe bottlenecks. Each B200 GPU in the system incorporates 192 GB of HBM3e memory with 8 TB/s of memory bandwidth per GPU, enabling the platform to hold and serve extremely large model parameter sets entirely in high-bandwidth memory without offloading.
Designed explicitly for AI factory deployments, the GB200 NVL72 targets organizations running frontier large language model training, multi-trillion-parameter model inference, and national-scale AI infrastructure. The rack integrates liquid cooling as the standard thermal solution to manage the platform's power envelope in hyperscale and enterprise data center environments. Its NVLink Switch topology allows every GPU in the rack to communicate directly with every other GPU at full bandwidth, making it the preferred foundation for distributed deep learning jobs that cannot tolerate the latency or bandwidth constraints of traditional scale-out networking alone.
| Manufacturer | NVIDIA |
| Manufacturer Part Number | 935-23B00-0000-000 |
| Platform Name | NVIDIA GB200 NVL72 |
| Form Factor | Rack-scale (full rack) |
| GPUs per Rack | 72 × NVIDIA Blackwell B200 GPUs |
| CPUs per Rack | 36 × NVIDIA Grace CPUs (ARM Neoverse V2-based) |
| GPU Architecture | NVIDIA Blackwell |
| GPU Memory per B200 | 192 GB HBM3e |
| Total GPU Memory (Rack) | 13.5 TB HBM3e |
| GPU Memory Bandwidth per B200 | 8 TB/s |
| Total GPU Memory Bandwidth (Rack) | 576 TB/s |
| AI Compute (FP4, Rack) | 1.4 exaflops |
| AI Compute (FP8, Rack) | 720 petaflops |
| CPU-GPU Interconnect | NVLink-C2C at 900 GB/s per Grace Blackwell Superchip (bidirectional) |
| Intra-Rack GPU Interconnect | NVIDIA NVLink 5 (fifth generation), 1.8 TB/s all-to-all aggregate bandwidth |
| NVLink Switch Generation | NVLink Switch (fifth generation) |
| External Fabric Support | NVIDIA InfiniBand and Ethernet (for multi-rack scale-out) |
| Thermal Solution | Liquid cooling (required) |
| Transformer Engine Generation | Second generation (native FP4, FP6, FP8, BF16, FP16, FP32, TF32) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | NVID-93523B000000000 |
| Part Number | 935-23B00-0000-000 |
| Condition | New |
| Manufacturer Part Number | 935-23B00-0000-000 |
| Platform Name | NVIDIA GB200 NVL72 |
| Form Factor | Rack-scale (full rack) |
| GPUs per Rack | 72 × NVIDIA Blackwell B200 GPUs |
| CPUs per Rack | 36 × NVIDIA Grace CPUs (ARM Neoverse V2-based) |
| GPU Architecture | NVIDIA Blackwell |
| GPU Memory per B200 | 192 GB HBM3e |
| Total GPU Memory (Rack) | 13.5 TB HBM3e |
| GPU Memory Bandwidth per B200 | 8 TB/s |
| Total GPU Memory Bandwidth (Rack) | 576 TB/s |
| AI Compute (FP4, Rack) | 1.4 exaflops |
| AI Compute (FP8, Rack) | 720 petaflops |
| CPU-GPU Interconnect | NVLink-C2C at 900 GB/s per Grace Blackwell Superchip (bidirectional) |
| Intra-Rack GPU Interconnect | NVIDIA NVLink 5 (fifth generation), 1.8 TB/s all-to-all aggregate bandwidth |
| NVLink Switch Generation | NVLink Switch (fifth generation) |
| External Fabric Support | NVIDIA InfiniBand and Ethernet (for multi-rack scale-out) |
| Thermal Solution | Liquid cooling (required) |
| Transformer Engine Generation | Second generation (native FP4, FP6, FP8, BF16, FP16, FP32, TF32) |
The NVIDIA GB200 NVL72 Grace Blackwell Superchip accelerates AI/ML training, inference, scientific HPC, and virtualization (vGPU) workloads. Typical deployments include LLM training clusters, computer-vision pipelines, financial risk modeling, and rendering farms.
Key specifications for the NVIDIA GB200 NVL72 Grace Blackwell Superchip: new condition; manufacturer NVIDIA; manufacturer part number 935-23B00-0000-000; platform name NVIDIA GB200 NVL72; form factor Rack-scale (full rack); gpus per rack 72 × NVIDIA Blackwell B200 GPUs; cpus per rack 36 × NVIDIA Grace CPUs (ARM Neoverse V2-based). Manufacturer part number 935-23B00-0000-000. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.
The NVIDIA GB200 NVL72 Grace Blackwell Superchip requires a PCIe Gen4 or Gen5 x16 slot, server power adequate for the card's TDP, and CUDA/ROCm driver support in your hypervisor or bare-metal OS. Sales engineering will confirm chassis fit (1U/2U/4U), PCIe lane count, and PSU headroom before quoting.