NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e)

NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e)

Brand: NVIDIA | Category: Servers

SKU: 935-24287-0070-000 | Part #: 935-24287-0070-000 | MPN: 935-24287-0070-000

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e)

The NVIDIA DGX H200 is a purpose-built AI supercomputing system integrating eight NVIDIA H200 Tensor Core GPUs interconnected via fifth-generation NVLink, delivering a combined 1.1 TB of HBM3e memory with 3.35 TB/s of aggregate memory bandwidth. Built on the same Hopper GPU architecture as the H100, the H200 advances the platform with significantly higher memory capacity and bandwidth, enabling larger models to reside entirely in GPU memory and dramatically reducing data movement bottlenecks during training and inference. The system is engineered for maximum out-of-the-box performance, arriving as a fully integrated, validated, and thermally optimized node ready for immediate deployment into enterprise data centers and AI infrastructure environments.

At its core, each H200 GPU delivers up to 3.958 TFLOPS of FP8 Transformer Engine performance and 989 PFLOPS at FP8 precision per GPU, with the full eight-GPU system delivering leadership-class throughput for large language model training, fine-tuning, and inference serving. NVLink Switch fabric provides 900 GB/s of bidirectional bandwidth between GPUs, treating all eight accelerators as a unified memory and compute pool. The DGX H200 includes dual Intel Xeon Scalable processors, 2 TB of system DRAM, high-throughput NVMe storage, and eight ConnectX-7 400GbE/NDR InfiniBand network adapters supporting scale-out cluster configurations.

The DGX H200 is targeted at organizations running frontier-scale generative AI workloads, scientific simulation, drug discovery pipelines, and real-time inference at enterprise scale. NVIDIA ships the system with DGX OS (an optimized Ubuntu-based software stack), the Base Command software suite for cluster orchestration, and access to the NGC container registry providing pre-optimized frameworks including PyTorch, TensorFlow, and RAPIDS. Its high memory capacity makes it particularly suited for training and serving models with hundreds of billions of parameters without model parallelism overheads that constrain lower-memory systems.

Ideal for

  • Training and fine-tuning large language models (LLMs) with 70B–700B+ parameters using full model weights resident in unified GPU HBM3e memory
  • High-throughput generative AI inference serving for enterprise conversational AI, code generation, and multimodal applications requiring low-latency responses at scale
  • Computational drug discovery and molecular dynamics simulation leveraging GPU-accelerated physics engines and deep learning-based protein structure prediction
  • Large-scale computer vision and multimodal model development for autonomous systems, medical imaging analysis, and industrial inspection pipelines
  • Climate and weather modeling, seismic processing, and scientific HPC workloads requiring both FP64 double-precision accuracy and AI-accelerated surrogate modeling
  • Retrieval-augmented generation (RAG) and enterprise AI platform infrastructure where high memory bandwidth reduces embedding and vector search latency at production scale

Technical specifications

ManufacturerNVIDIA
Manufacturer Part Number935-24287-0070-000
GPU8x NVIDIA H200 Tensor Core GPU
GPU ArchitectureNVIDIA Hopper (GH100)
Total GPU Memory1.1 TB HBM3e
GPU Memory per GPU141 GB HBM3e
Aggregate GPU Memory Bandwidth3.35 TB/s
FP8 Tensor Core Performance (per GPU)3,958 TFLOPS (with Sparsity)
FP16 / BF16 Tensor Core Performance (per GPU)1,979 TFLOPS (with Sparsity)
FP64 Tensor Core Performance (per GPU)67 TFLOPS
GPU InterconnectNVLink 4.0, 900 GB/s bidirectional total bandwidth (all-to-all)
CPUDual Intel Xeon Platinum 8480C (Sapphire Rapids), 60 cores each
System Memory2 TB DDR5
Internal Storage8x 3.84 TB NVMe SSD (30.72 TB raw)
Network Adapters8x NVIDIA ConnectX-7 VPI (400GbE / NDR 400Gb/s InfiniBand)
Out-of-Band ManagementDedicated BMC with NVIDIA IPMI / Redfish support
Form Factor6U Rackmount
Power SupplyRedundant 10 kW (N+1)
Operating SystemDGX OS (Ubuntu-based, NVIDIA optimized)
Software StackNVIDIA Base Command, NGC Container Registry, CUDA 12.x, cuDNN, NCCL

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandNVIDIA
CategoryServers
SKU935-24287-0070-000
Part Number935-24287-0070-000
ConditionNew
Manufacturer Part Number935-24287-0070-000
GPU8x NVIDIA H200 Tensor Core GPU
GPU ArchitectureNVIDIA Hopper (GH100)
Total GPU Memory1.1 TB HBM3e
GPU Memory per GPU141 GB HBM3e
Aggregate GPU Memory Bandwidth3.35 TB/s
FP8 Tensor Core Performance (per GPU)3,958 TFLOPS (with Sparsity)
FP16 / BF16 Tensor Core Performance (per GPU)1,979 TFLOPS (with Sparsity)
FP64 Tensor Core Performance (per GPU)67 TFLOPS
GPU InterconnectNVLink 4.0, 900 GB/s bidirectional total bandwidth (all-to-all)
CPUDual Intel Xeon Platinum 8480C (Sapphire Rapids), 60 cores each
System Memory2 TB DDR5
Internal Storage8x 3.84 TB NVMe SSD (30.72 TB raw)
Network Adapters8x NVIDIA ConnectX-7 VPI (400GbE / NDR 400Gb/s InfiniBand)
Out-of-Band ManagementDedicated BMC with NVIDIA IPMI / Redfish support
Form Factor6U Rackmount
Power SupplyRedundant 10 kW (N+1)
Operating SystemDGX OS (Ubuntu-based, NVIDIA optimized)
Software StackNVIDIA Base Command, NGC Container Registry, CUDA 12.x, cuDNN, NCCL

Frequently Asked Questions about NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e)

What does the NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e) do?

The NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e) is built for enterprise data-center workloads — virtualization (VMware, Proxmox, Nutanix), private cloud, database hosting, and AI/ML training. It fits standard EIA-310 server racks and supports redundant PSUs and hot-swap drives common in production environments.

What are the headline specs of the NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e)?

Key specifications for the NVIDIA DGX H200 AI System (8x H200 Tensor Core GPU, 1.1TB HBM3e): new condition; manufacturer NVIDIA; manufacturer part number 935-24287-0070-000; gpu 8x NVIDIA H200 Tensor Core GPU; gpu architecture NVIDIA Hopper (GH100); total gpu memory 1.1 TB HBM3e; gpu memory per gpu 141 GB HBM3e. Manufacturer part number 935-24287-0070-000. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.