HPE NVIDIA H200 141GB SXM GPU Module

HPE NVIDIA H200 141GB SXM GPU Module

Brand: NVIDIA | Category: GPUs

SKU: S2L01A | Part #: S2L01A | MPN: S2L01A

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the HPE NVIDIA H200 141GB SXM GPU Module

The HPE NVIDIA H200 141GB SXM GPU Module is a cutting-edge GPU accelerator featuring NVIDIA's Hopper architecture with 141GB of HBM3e memory, delivering exceptional performance for large-scale AI model training, inference, and high-performance computing workloads. The module integrates seamlessly into HPE ProLiant XL675d v11 and compatible systems via SXM socket connectivity, enabling enterprises to build cohesive AI/HPC clusters with unified memory bandwidth and low-latency GPU-to-GPU interconnects.

With 18,176 CUDA cores and specialized Tensor engines optimized for FP8, FP32, and mixed-precision operations, the H200 accelerates transformer model training, retrieval-augmented generation (RAG), and complex scientific simulations. The 141GB HBM3e capacity enables in-GPU processing of larger working datasets and attention mechanisms, reducing host-device memory transfers and bottlenecks inherent in smaller-memory GPU configurations.

The module operates within a standardized SXM form factor, supporting up to 8 GPUs per server chassis for distributed training and inference at scale. Enterprise IT environments leverage the H200 to address emerging demands in generative AI, foundation model fine-tuning, and computational research without requiring architectural redesigns of existing GPU-capable infrastructure.

Ideal for

  • Large language model (LLM) training and instruction-tuning on datasets exceeding 100B tokens
  • Production inference serving for multimodal and retrieval-augmented generation (RAG) applications
  • Molecular dynamics simulation and computational chemistry for drug discovery and materials science
  • High-resolution climate modeling and scientific visualization in research institutions
  • Graph neural network training for recommendation systems and fraud detection at enterprise scale
  • Real-time video processing and computer vision pipelines for autonomous systems and surveillance

Technical specifications

ManufacturerNVIDIA
BrandHPE NVIDIA
ModelH200
Manufacturer Part NumberS2L01A
GPU Memory141GB HBM3e
Memory Bandwidth4.8 TB/s
CUDA Cores18,176
Tensor Float 32 (TF32) Performance1.457 TFLOPS
BFLOAT16 Performance2.914 TFLOPS
FP8 Performance2.914 TFLOPS
ArchitectureNVIDIA Hopper
Form FactorSXM (Socket Mountable eXpansion)
InterconnectNVIDIA NVLink 4.0 (900 GB/s per link)
Max GPUs per Server8
Power Consumption700W
Compatible ServersHPE ProLiant XL675d v11, XL675d v10 (with BIOS updates)
CoolingLiquid-cooled module with integrated cold plate
Release TimelineQ1 2025

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandNVIDIA
CategoryGPUs
SKUS2L01A
Part NumberS2L01A
ConditionNew
ModelH200
Manufacturer Part NumberS2L01A
GPU Memory141GB HBM3e
Memory Bandwidth4.8 TB/s
CUDA Cores18,176
Tensor Float 32 (TF32) Performance1.457 TFLOPS
BFLOAT16 Performance2.914 TFLOPS
FP8 Performance2.914 TFLOPS
ArchitectureNVIDIA Hopper
Form FactorSXM (Socket Mountable eXpansion)
InterconnectNVIDIA NVLink 4.0 (900 GB/s per link)
Max GPUs per Server8
Power Consumption700W
Compatible ServersHPE ProLiant XL675d v11, XL675d v10 (with BIOS updates)
CoolingLiquid-cooled module with integrated cold plate
Release TimelineQ1 2025