Omnixon Global
Supermicro NVIDIA GB200 NVL72 48U Rack Solution

Supermicro NVIDIA GB200 NVL72 48U Rack Solution

Brand: Supermicro | Category: GPUs

SKU: ARS-111GL-NHR (MGX NVL72) | Part #: SRS-GB200-NVL72 | MPN: SRS-GB200-NVL72

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Supermicro NVIDIA GB200 NVL72 48U Rack Solution

The Supermicro NVIDIA GB200 NVL72 48U Rack Solution (SRS-GB200-NVL72) is a rack-scale AI supercomputing system built on NVIDIA's Blackwell architecture, integrating 36 Grace CPU modules and 72 Blackwell B200 GPU dies within a single 48U rack enclosure. Each GB200 superchip pairs two B200 GPUs with one Grace Arm-based CPU via NVLink-C2C interconnect, delivering a unified, high-bandwidth memory and compute fabric. The system leverages NVLink Switch technology to unite all 72 GPUs into a single high-speed NVLink domain at 1.8 TB/s per GPU of NVLink bandwidth, enabling the entire rack to operate as one logical GPU for large-scale model training and inference.

The GB200 NVL72 rack is engineered for the most demanding frontier AI workloads, including trillion-parameter large language model (LLM) training, mixture-of-experts (MoE) inference at scale, and scientific simulation. The Grace CPU subsystem contributes 480 GB of LPDDR5X memory per rack accessible at CPU-class bandwidth, while each B200 GPU delivers 192 GB of HBM3e memory, resulting in approximately 13.8 TB of total HBM3e capacity across the rack. Supermicro's MGX mechanical design framework enables the dense integration of compute, networking, and power delivery into the 48U form factor with liquid cooling as the primary thermal solution.

The system is designed for direct liquid cooling (DLC) to support its exceptional thermal load, making it suited for modern data centers equipped with liquid cooling infrastructure. High-speed networking is supported through NVIDIA Quantum-X800 InfiniBand or Spectrum-X800 Ethernet fabric connectivity, enabling multi-rack cluster scaling for hyperscale AI training clusters. The SRS-GB200-NVL72 is positioned for enterprise data centers, national AI labs, cloud service providers, and large-scale HPC facilities across UAE, GCC, EMEA, and APAC regions seeking to deploy next-generation AI infrastructure.

Ideal for

  • Frontier large language model (LLM) pre-training and fine-tuning at trillion-parameter scale, leveraging the unified 72-GPU NVLink domain as a single logical compute unit
  • High-throughput generative AI inference serving for production enterprise deployments requiring maximum tokens-per-second throughput with reduced latency
  • Scientific high-performance computing (HPC) simulations including climate modeling, molecular dynamics, and computational fluid dynamics benefiting from the combined Grace CPU and Blackwell GPU architecture
  • Mixture-of-experts (MoE) model training and inference where the high-capacity HBM3e memory pool across 72 GPUs accommodates extremely large sparse model states
  • Multi-modal AI workload execution combining vision, language, and structured data processing within a single rack-scale unified memory and compute environment
  • Sovereign AI and national-scale AI infrastructure deployments in GCC, EMEA, and APAC data centers requiring certified, rack-integrated GPU systems with enterprise-grade support

Technical specifications

ManufacturerSupermicro
Manufacturer Part NumberSRS-GB200-NVL72
Product FamilySupermicro MGX NVL72 GB200 Rack-Scale System
GPU ArchitectureNVIDIA Blackwell
GPU ModelNVIDIA B200 (GB200 NVL72 configuration)
Total GPUs per Rack72 x NVIDIA B200 GPU dies
Total CPU Modules per Rack36 x NVIDIA Grace Arm-based CPU modules
GPU InterconnectNVLink 5, 1.8 TB/s bidirectional per GPU
NVLink Domain72-GPU unified NVLink domain via NVLink Switch
CPU-to-GPU InterconnectNVLink-C2C (chip-to-chip), 900 GB/s bidirectional per GB200 superchip
HBM Memory per GPU192 GB HBM3e
Total HBM3e Capacity (Rack)Approximately 13.8 TB
CPU Memory per RackApproximately 480 GB LPDDR5X (Grace CPU memory)
Rack Form Factor48U
CoolingDirect Liquid Cooling (DLC)
Networking OptionsNVIDIA Quantum-X800 InfiniBand / Spectrum-X800 Ethernet
Platform DesignNVIDIA MGX mechanical framework
Target DeploymentData center, HPC, and AI supercomputing facilities
Region AvailabilityUAE, GCC, EMEA, APAC

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandSupermicro
CategoryGPUs
SKUARS-111GL-NHR (MGX NVL72)
Part NumberSRS-GB200-NVL72
ConditionNew
Manufacturer Part NumberSRS-GB200-NVL72
Product FamilySupermicro MGX NVL72 GB200 Rack-Scale System
GPU ArchitectureNVIDIA Blackwell
GPU ModelNVIDIA B200 (GB200 NVL72 configuration)
Total GPUs per Rack72 x NVIDIA B200 GPU dies
Total CPU Modules per Rack36 x NVIDIA Grace Arm-based CPU modules
GPU InterconnectNVLink 5, 1.8 TB/s bidirectional per GPU
NVLink Domain72-GPU unified NVLink domain via NVLink Switch
CPU-to-GPU InterconnectNVLink-C2C (chip-to-chip), 900 GB/s bidirectional per GB200 superchip
HBM Memory per GPU192 GB HBM3e
Total HBM3e Capacity (Rack)Approximately 13.8 TB
CPU Memory per RackApproximately 480 GB LPDDR5X (Grace CPU memory)
Rack Form Factor48U
CoolingDirect Liquid Cooling (DLC)
Networking OptionsNVIDIA Quantum-X800 InfiniBand / Spectrum-X800 Ethernet
Platform DesignNVIDIA MGX mechanical framework
Target DeploymentData center, HPC, and AI supercomputing facilities
Region AvailabilityUAE, GCC, EMEA, APAC

Frequently Asked Questions about Supermicro NVIDIA GB200 NVL72 48U Rack Solution

What server platforms accept the Supermicro MGX NVL72 GB200 NVL72 Rack-Scale System?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.