Brand: HPE | Category: GPUs
SKU: P57602-B21 | Part #: P57602-B21 | MPN: P57602-B21
Contact for Pricing — Request a Quote
The NVIDIA GH200 480GB NVL32 Grace Hopper Superchip for HPE Cray XD670 (P57602-B21) is a purpose-built accelerated computing module that unifies an NVIDIA Grace CPU with an NVIDIA Hopper GPU on a single package, connected via NVLink-C2C with 900 GB/s of chip-to-chip bandwidth. The GH200 Superchip delivers 96GB of HBM3 GPU memory paired with 480GB of LPDDR5X CPU memory in the NVL32 configuration, enabling exceptionally large model footprints to reside entirely in high-bandwidth memory across a 32-node NVLink domain without reliance on slower PCIe interconnects.
Designed natively for the HPE Cray XD670 platform, this module integrates seamlessly into HPE's liquid-cooled, high-density accelerated compute architecture. The Hopper GPU die incorporates fourth-generation Tensor Cores with FP8 precision support, a Transformer Engine for dynamic mixed-precision inference and training, and Confidential Computing capabilities via hardware-level TEE isolation. The NVL32 multi-node NVLink configuration allows up to 32 GH200 Superchips to share a unified 57.6 TB/s aggregate NVLink bandwidth fabric, enabling tight coupling of GPU and CPU memory across nodes at scales not achievable with conventional interconnects.
This accelerator is engineered for the most demanding generative AI, large language model training and inference, scientific simulation, and high-performance computing workloads in enterprise data centers. Its combination of massive coherent memory capacity, extreme interconnect bandwidth, and Hopper-generation compute throughput positions the GH200 NVL32 as a foundational building block for next-generation AI infrastructure deployed across hyperscale, national laboratory, and enterprise research environments.
| Manufacturer | HPE |
| Manufacturer Part Number | P57602-B21 |
| Product Name | NVIDIA GH200 480GB NVL32 Grace Hopper Superchip for HPE Cray XD670 |
| GPU Architecture | NVIDIA Hopper (GH100) |
| CPU Architecture | NVIDIA Grace (72-core Arm Neoverse V2) |
| GPU Memory | 96 GB HBM3 |
| CPU Memory | 480 GB LPDDR5X |
| GPU Memory Bandwidth | 4.0 TB/s |
| CPU Memory Bandwidth | 512 GB/s |
| NVLink-C2C Chip-to-Chip Bandwidth | 900 GB/s (bidirectional) |
| NVLink Domain Configuration | NVL32 (up to 32 GH200 Superchips per NVLink domain) |
| NVLink Domain Aggregate Bandwidth | 57.6 TB/s |
| Tensor Core Generation | 4th Generation (FP8, FP16, BF16, TF32, FP64) |
| Transformer Engine | Yes (dynamic FP8/FP16 mixed precision) |
| Confidential Computing | Yes (hardware-level Trusted Execution Environment) |
| Form Factor | OAM (OCP Accelerator Module), liquid-cooled |
| Compatible Platform | HPE Cray XD670 |
| Interconnect | NVLink-C2C (CPU-GPU), NVLink 4.0 (GPU-to-GPU across nodes) |
| PCIe Generation | PCIe Gen 5 |
| Cooling | Liquid cooling (direct liquid cooling via HPE Cray XD670 chassis) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P57602-B21 |
| Part Number | P57602-B21 |
| Condition | New |
| Manufacturer Part Number | P57602-B21 |
| Product Name | NVIDIA GH200 480GB NVL32 Grace Hopper Superchip for HPE Cray XD670 |
| GPU Architecture | NVIDIA Hopper (GH100) |
| CPU Architecture | NVIDIA Grace (72-core Arm Neoverse V2) |
| GPU Memory | 96 GB HBM3 |
| CPU Memory | 480 GB LPDDR5X |
| GPU Memory Bandwidth | 4.0 TB/s |
| CPU Memory Bandwidth | 512 GB/s |
| NVLink-C2C Chip-to-Chip Bandwidth | 900 GB/s (bidirectional) |
| NVLink Domain Configuration | NVL32 (up to 32 GH200 Superchips per NVLink domain) |
| NVLink Domain Aggregate Bandwidth | 57.6 TB/s |
| Tensor Core Generation | 4th Generation (FP8, FP16, BF16, TF32, FP64) |
| Transformer Engine | Yes (dynamic FP8/FP16 mixed precision) |
| Confidential Computing | Yes (hardware-level Trusted Execution Environment) |
| Form Factor | OAM (OCP Accelerator Module), liquid-cooled |
| Compatible Platform | HPE Cray XD670 |
| Interconnect | NVLink-C2C (CPU-GPU), NVLink 4.0 (GPU-to-GPU across nodes) |
| PCIe Generation | PCIe Gen 5 |
| Cooling | Liquid cooling (direct liquid cooling via HPE Cray XD670 chassis) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.