Brand: HPE | Category: GPUs
SKU: P65891-B21 | Part #: P65891-B21 | MPN: P65891-B21
Contact for Pricing — Request a Quote
The HPE NVIDIA H200 SXM5 141GB GPU Computing Module (P65891-B21) is built on NVIDIA's Hopper architecture and delivers a major leap in high-bandwidth memory capacity over its predecessor, equipping enterprises with 141GB of HBM3e memory to handle the most memory-intensive AI and HPC workloads at scale. Engineered for integration into HPE's AI-optimized server platforms, this module connects via the SXM5 form factor, enabling the high-speed NVLink and NVSwitch interconnect fabric that sustains extreme bandwidth between GPUs in multi-GPU configurations.
At the core of the H200 is NVIDIA's fourth-generation Transformer Engine, which accelerates large language model (LLM) training and inference with support for FP8, FP16, BF16, TF32, and INT8 precision formats. The expanded 141GB HBM3e memory pool — paired with substantial memory bandwidth — directly addresses the bottleneck that limits serving very large foundation models and running long-context inference, making it a critical enabler for next-generation generative AI deployments. The module also retains full Hopper-generation capabilities including second-generation Multi-Instance GPU (MIG) technology for granular workload partitioning.
As an HPE-branded computing module, P65891-B21 is validated and integrated within HPE's ecosystem of AI infrastructure solutions, ensuring compatibility with HPE system management tools, firmware update pipelines, and data center thermal and power standards. This makes it well suited for enterprises in regulated industries or those requiring cohesive, vendor-validated AI infrastructure stacks across UAE, GCC, EMEA, and APAC regions.
| Manufacturer | HPE |
| Manufacturer Part Number | P65891-B21 |
| GPU Model | NVIDIA H200 SXM5 |
| Architecture | NVIDIA Hopper (GH100) |
| Form Factor | SXM5 |
| GPU Memory | 141 GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| FP8 Tensor Core Performance | 3,958 TFLOPS |
| FP16 / BF16 Tensor Core Performance | 1,979 TFLOPS |
| TF32 Tensor Core Performance | 989 TFLOPS |
| FP64 Tensor Core Performance | 67 TFLOPS |
| GPU Interconnect | NVLink 4.0 (900 GB/s bidirectional per GPU) |
| NVLink Bandwidth | 900 GB/s |
| PCIe Interface | PCIe Gen5 |
| Multi-Instance GPU (MIG) | Yes — up to 7 MIG instances |
| Transformer Engine Generation | 4th Generation |
| Thermal Design Power (TDP) | 700 W |
| Supported Precision Formats | FP8, FP16, BF16, TF32, FP32, FP64, INT8 |
| ECC Memory Support | Yes |
| Platform Compatibility | HPE AI-optimized server platforms supporting SXM5 module configuration |
| Category | GPU Computing Module |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P65891-B21 |
| Part Number | P65891-B21 |
| Condition | New |
| Manufacturer Part Number | P65891-B21 |
| GPU Model | NVIDIA H200 SXM5 |
| Architecture | NVIDIA Hopper (GH100) |
| Form Factor | SXM5 |
| GPU Memory | 141 GB HBM3e |
| Memory Bandwidth | 4.8 TB/s |
| FP8 Tensor Core Performance | 3,958 TFLOPS |
| FP16 / BF16 Tensor Core Performance | 1,979 TFLOPS |
| TF32 Tensor Core Performance | 989 TFLOPS |
| FP64 Tensor Core Performance | 67 TFLOPS |
| GPU Interconnect | NVLink 4.0 (900 GB/s bidirectional per GPU) |
| NVLink Bandwidth | 900 GB/s |
| PCIe Interface | PCIe Gen5 |
| Multi-Instance GPU (MIG) | Yes — up to 7 MIG instances |
| Transformer Engine Generation | 4th Generation |
| Thermal Design Power (TDP) | 700 W |
| Supported Precision Formats | FP8, FP16, BF16, TF32, FP32, FP64, INT8 |
| ECC Memory Support | Yes |
| Platform Compatibility | HPE AI-optimized server platforms supporting SXM5 module configuration |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.