HPE NVIDIA H200 SXM5 141GB GPU Computing Module

HPE NVIDIA H200 SXM5 141GB GPU Computing Module

Brand: HPE | Category: GPUs

SKU: P65891-B21 | Part #: P65891-B21 | MPN: P65891-B21

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the HPE NVIDIA H200 SXM5 141GB GPU Computing Module

The HPE NVIDIA H200 SXM5 141GB GPU Computing Module (P65891-B21) is built on NVIDIA's Hopper architecture and delivers a major leap in high-bandwidth memory capacity over its predecessor, equipping enterprises with 141GB of HBM3e memory to handle the most memory-intensive AI and HPC workloads at scale. Engineered for integration into HPE's AI-optimized server platforms, this module connects via the SXM5 form factor, enabling the high-speed NVLink and NVSwitch interconnect fabric that sustains extreme bandwidth between GPUs in multi-GPU configurations.

At the core of the H200 is NVIDIA's fourth-generation Transformer Engine, which accelerates large language model (LLM) training and inference with support for FP8, FP16, BF16, TF32, and INT8 precision formats. The expanded 141GB HBM3e memory pool — paired with substantial memory bandwidth — directly addresses the bottleneck that limits serving very large foundation models and running long-context inference, making it a critical enabler for next-generation generative AI deployments. The module also retains full Hopper-generation capabilities including second-generation Multi-Instance GPU (MIG) technology for granular workload partitioning.

As an HPE-branded computing module, P65891-B21 is validated and integrated within HPE's ecosystem of AI infrastructure solutions, ensuring compatibility with HPE system management tools, firmware update pipelines, and data center thermal and power standards. This makes it well suited for enterprises in regulated industries or those requiring cohesive, vendor-validated AI infrastructure stacks across UAE, GCC, EMEA, and APAC regions.

Ideal for

  • Training and fine-tuning large language models (LLMs) and multimodal foundation models requiring high memory capacity per GPU
  • High-throughput generative AI inference serving for production deployments of models exceeding 70 billion parameters
  • High-performance computing (HPC) simulations in life sciences, computational fluid dynamics, and climate modeling that demand large in-GPU memory footprints
  • Enterprise AI data pipeline acceleration including embedding generation, vector search preprocessing, and retrieval-augmented generation (RAG) workloads
  • Multi-tenant GPU partitioning via MIG for cloud-style shared AI infrastructure within private data centers
  • Scientific research and national laboratory-scale workloads requiring FP64 double-precision compute alongside AI acceleration

Technical specifications

ManufacturerHPE
Manufacturer Part NumberP65891-B21
GPU ModelNVIDIA H200 SXM5
ArchitectureNVIDIA Hopper (GH100)
Form FactorSXM5
GPU Memory141 GB HBM3e
Memory Bandwidth4.8 TB/s
FP8 Tensor Core Performance3,958 TFLOPS
FP16 / BF16 Tensor Core Performance1,979 TFLOPS
TF32 Tensor Core Performance989 TFLOPS
FP64 Tensor Core Performance67 TFLOPS
GPU InterconnectNVLink 4.0 (900 GB/s bidirectional per GPU)
NVLink Bandwidth900 GB/s
PCIe InterfacePCIe Gen5
Multi-Instance GPU (MIG)Yes — up to 7 MIG instances
Transformer Engine Generation4th Generation
Thermal Design Power (TDP)700 W
Supported Precision FormatsFP8, FP16, BF16, TF32, FP32, FP64, INT8
ECC Memory SupportYes
Platform CompatibilityHPE AI-optimized server platforms supporting SXM5 module configuration
CategoryGPU Computing Module

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandHPE
CategoryGPUs
SKUP65891-B21
Part NumberP65891-B21
ConditionNew
Manufacturer Part NumberP65891-B21
GPU ModelNVIDIA H200 SXM5
ArchitectureNVIDIA Hopper (GH100)
Form FactorSXM5
GPU Memory141 GB HBM3e
Memory Bandwidth4.8 TB/s
FP8 Tensor Core Performance3,958 TFLOPS
FP16 / BF16 Tensor Core Performance1,979 TFLOPS
TF32 Tensor Core Performance989 TFLOPS
FP64 Tensor Core Performance67 TFLOPS
GPU InterconnectNVLink 4.0 (900 GB/s bidirectional per GPU)
NVLink Bandwidth900 GB/s
PCIe InterfacePCIe Gen5
Multi-Instance GPU (MIG)Yes — up to 7 MIG instances
Transformer Engine Generation4th Generation
Thermal Design Power (TDP)700 W
Supported Precision FormatsFP8, FP16, BF16, TF32, FP32, FP64, INT8
ECC Memory SupportYes
Platform CompatibilityHPE AI-optimized server platforms supporting SXM5 module configuration

Frequently Asked Questions about HPE NVIDIA H200 SXM5 141GB GPU Computing Module

What server platforms accept the HPE NVIDIA H200 SXM5 141GB GPU Computing Module?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.