Brand: HPE | Category: GPUs
SKU: P57601-B21 | Part #: P57601-B21 | MPN: P57601-B21
Contact for Pricing — Request a Quote
The NVIDIA H200 141GB NVL PCIe Gen5 GPU for HPE ProLiant DL380 Gen11 (P57601-B21) is built on NVIDIA's Hopper architecture and represents the highest-capacity HBM3e-based accelerator available in PCIe form factor. With 141GB of HBM3e memory and a 4.8TB/s memory bandwidth, this GPU delivers exceptional throughput for large-scale AI inference, LLM training, and high-performance computing workloads. The NVL designation indicates the full 141GB memory configuration, making it purpose-fit for workloads that require holding entire large language models or massive scientific datasets entirely in GPU memory.
Designed for HPE ProLiant DL380 Gen11 server platforms, this GPU is validated and factory-integrated by HPE as part of their Compute portfolio, ensuring full hardware and firmware compatibility. The PCIe Gen5 host interface provides double the bandwidth of PCIe Gen4, minimizing host-to-device data transfer bottlenecks in CPU-GPU heterogeneous computing pipelines. The H200 NVL features fourth-generation Tensor Cores and second-generation Transformer Engines, delivering significantly improved FP8 and FP16 throughput compared to its A100 predecessor, and is fully compatible with NVIDIA's CUDA ecosystem, cuDNN, and TensorRT inference frameworks.
As an HPE-branded and validated component (P57601-B21), this accelerator integrates with HPE's iLO management infrastructure and is supported through HPE's enterprise support channels, making it a production-grade choice for data centers requiring vendor-managed lifecycle support. It is suited for organizations in UAE, GCC, EMEA, and APAC regions building out AI infrastructure, HPC clusters, and large-scale data analytics platforms on HPE ProLiant Gen11 hardware.
| Manufacturer | HPE |
| Manufacturer Part Number | P57601-B21 |
| GPU Model | NVIDIA H200 NVL |
| Architecture | NVIDIA Hopper (GH100) |
| Memory Capacity | 141 GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| Host Interface | PCIe Gen5 x16 |
| Form Factor | Dual-slot, full-height full-length (FHFL) |
| Tensor Core Generation | 4th Generation |
| Transformer Engine Generation | 2nd Generation |
| FP8 Tensor Core Performance | 3,958 TFLOPS |
| FP16 Tensor Core Performance | 1,979 TFLOPS |
| BF16 Tensor Core Performance | 1,979 TFLOPS |
| FP64 Tensor Core Performance | 67 TFLOPS |
| TDP (Thermal Design Power) | 700W |
| NVLink Support | No (PCIe variant; NVLink requires SXM form factor) |
| Compatible Server | HPE ProLiant DL380 Gen11 |
| CUDA Compute Capability | 9.0 |
| ECC Memory Support | Yes |
| Confidential Computing | Yes (NVIDIA Hopper Confidential Computing) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P57601-B21 |
| Part Number | P57601-B21 |
| Condition | New |
| Manufacturer Part Number | P57601-B21 |
| GPU Model | NVIDIA H200 NVL |
| Architecture | NVIDIA Hopper (GH100) |
| Memory Capacity | 141 GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| Host Interface | PCIe Gen5 x16 |
| Form Factor | Dual-slot, full-height full-length (FHFL) |
| Tensor Core Generation | 4th Generation |
| Transformer Engine Generation | 2nd Generation |
| FP8 Tensor Core Performance | 3,958 TFLOPS |
| FP16 Tensor Core Performance | 1,979 TFLOPS |
| BF16 Tensor Core Performance | 1,979 TFLOPS |
| FP64 Tensor Core Performance | 67 TFLOPS |
| TDP (Thermal Design Power) | 700W |
| NVLink Support | No (PCIe variant; NVLink requires SXM form factor) |
| Compatible Server | HPE ProLiant DL380 Gen11 |
| CUDA Compute Capability | 9.0 |
| ECC Memory Support | Yes |
| Confidential Computing | Yes (NVIDIA Hopper Confidential Computing) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.