Dell PowerEdge XE9712 with NVIDIA GB200 NVL72 (Grace Blackwell)

Dell PowerEdge XE9712 with NVIDIA GB200 NVL72 (Grace Blackwell)

Brand: Dell | Category: GPUs

SKU: PowerEdge XE9712 | Part #: PowerEdge XE9712 | MPN: PowerEdge XE9712

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Dell PowerEdge XE9712 with NVIDIA GB200 NVL72 (Grace Blackwell)

The Dell PowerEdge XE9712 is a purpose-built, rack-scale AI infrastructure platform engineered around the NVIDIA GB200 NVL72 Grace Blackwell Superchip architecture. The system integrates 36 NVIDIA GB200 Superchips — each pairing one NVIDIA Grace CPU with two NVIDIA Blackwell B200 GPUs — delivering 72 Blackwell GPUs and 36 Grace CPUs within a single NVLink-interconnected rack domain. The GB200 NVL72 configuration provides up to 1.4 exaflops of AI compute at FP4 precision, with the 5th-generation NVLink fabric delivering 130 TB/s of aggregate GPU-to-GPU bandwidth across the entire 72-GPU pool, enabling the system to function as a unified, tightly coupled compute fabric rather than a collection of discrete nodes.

The XE9712 is designed around a direct liquid cooling (DLC) architecture to manage the extreme thermal output of the GB200 NVL72 rack. Each Blackwell B200 GPU features 192 GB of HBM3e memory, yielding a total HBM3e capacity of approximately 13.8 TB across the 72-GPU rack, with aggregate memory bandwidth exceeding 1.1 PB/s. The Grace CPU subsystem contributes additional LPDDR5X system memory per Superchip, connected to the GPU dies via NVIDIA's high-bandwidth NVLink-C2C chip-to-chip interconnect at 900 GB/s bidirectional bandwidth per Superchip, eliminating the traditional PCIe bottleneck between CPU and GPU.

As an enterprise-grade platform, the PowerEdge XE9712 is positioned for organizations deploying frontier-scale AI training, large language model inference, and high-performance simulation workloads at data center scale. Dell's integration of the NVIDIA MGX architecture within the XE9712 ensures compatibility with NVIDIA's full software ecosystem — including CUDA, NCCL, NVLink Switch System support, and NVIDIA AI Enterprise — while Dell's OpenManage systems management and IDRAC9 with Redfish API provide enterprise lifecycle management, telemetry, and operational visibility consistent with the broader PowerEdge portfolio.

Ideal for

  • Training and fine-tuning frontier large language models (LLMs) and multimodal foundation models requiring tightly coupled, high-bandwidth multi-GPU compute at scale
  • High-throughput generative AI inference serving for enterprise deployments demanding low-latency, high-concurrency token generation across extremely large model parameter counts
  • Accelerated scientific simulation and digital twin workloads in energy, life sciences, and climate modeling that require sustained FP64 and FP8 mixed-precision compute
  • Sovereign AI and national-scale AI infrastructure deployments where organizations require on-premises, liquid-cooled, rack-native GPU density under direct operational control
  • Recommender system training and real-time ranking inference at hyperscale, leveraging the XE9712's extreme memory bandwidth and NVLink all-to-all GPU communication topology
  • HPC and AI-converged workloads combining traditional supercomputing simulation with AI-driven analysis pipelines, utilizing the Grace CPU and Blackwell GPU within a unified memory-coherent architecture

Technical specifications

ManufacturerDell
Manufacturer Part NumberPowerEdge XE9712
Form FactorRack-scale system (NVL72 rack unit)
GPU ArchitectureNVIDIA Blackwell (5th generation)
GPU Configuration72 x NVIDIA B200 GPUs (via 36 GB200 Superchips)
CPU Configuration36 x NVIDIA Grace CPUs (ARM Neoverse V2, 72 cores each, integrated per Superchip)
AI Compute Performance (FP4)Up to 1.4 exaflops
AI Compute Performance (FP8)Up to 720 petaflops
GPU Memory per B200192 GB HBM3e
Total GPU HBM3e MemoryApproximately 13.8 TB across 72 GPUs
Aggregate GPU Memory BandwidthExceeds 1.1 PB/s
GPU Interconnect FabricNVIDIA NVLink 5 (5th generation NVLink Switch System)
NVLink Aggregate Bisection Bandwidth130 TB/s across 72-GPU NVL72 domain
CPU-to-GPU InterconnectNVIDIA NVLink-C2C, 900 GB/s bidirectional per Superchip
Cooling ArchitectureDirect Liquid Cooling (DLC)
Systems ManagementDell iDRAC9 with Redfish API, Dell OpenManage
Software EcosystemNVIDIA CUDA, NCCL, NVIDIA AI Enterprise, NVIDIA NIM compatible
Target SegmentEnterprise data center, sovereign AI, hyperscale AI training and inference

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandDell
CategoryGPUs
SKUPowerEdge XE9712
Part NumberPowerEdge XE9712
ConditionNew
Manufacturer Part NumberPowerEdge XE9712
Form FactorRack-scale system (NVL72 rack unit)
GPU ArchitectureNVIDIA Blackwell (5th generation)
GPU Configuration72 x NVIDIA B200 GPUs (via 36 GB200 Superchips)
CPU Configuration36 x NVIDIA Grace CPUs (ARM Neoverse V2, 72 cores each, integrated per Superchip)
AI Compute Performance (FP4)Up to 1.4 exaflops
AI Compute Performance (FP8)Up to 720 petaflops
GPU Memory per B200192 GB HBM3e
Total GPU HBM3e MemoryApproximately 13.8 TB across 72 GPUs
Aggregate GPU Memory BandwidthExceeds 1.1 PB/s
GPU Interconnect FabricNVIDIA NVLink 5 (5th generation NVLink Switch System)
NVLink Aggregate Bisection Bandwidth130 TB/s across 72-GPU NVL72 domain
CPU-to-GPU InterconnectNVIDIA NVLink-C2C, 900 GB/s bidirectional per Superchip
Cooling ArchitectureDirect Liquid Cooling (DLC)
Systems ManagementDell iDRAC9 with Redfish API, Dell OpenManage
Software EcosystemNVIDIA CUDA, NCCL, NVIDIA AI Enterprise, NVIDIA NIM compatible
Target SegmentEnterprise data center, sovereign AI, hyperscale AI training and inference

Frequently Asked Questions about Dell PowerEdge XE9712 with NVIDIA GB200 NVL72 (Grace Blackwell)

What server platforms accept the Dell PowerEdge XE9712 with NVIDIA GB200 NVL72 (Grace Blackwell)?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.