HPE Cray XD675 AI Accelerator Server

HPE Cray XD675 AI Accelerator Server

Brand: HPE | Category: Servers

SKU: HPE-R9G65A | Part #: R9G65A | MPN: R9G65A

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the HPE Cray XD675 AI Accelerator Server

The HPE Cray XD675 is a purpose-built 4U AI accelerator server designed within the Open Compute Project (OCP) Data Center Ready form factor, engineered to support the most demanding large language model (LLM) training, fine-tuning, and high-performance computing workloads. At its core, the XD675 accommodates up to eight NVIDIA H200 SXM5 Tensor Core GPUs interconnected via fifth-generation NVLink, delivering a combined GPU memory capacity of up to 1.12 TB of HBM3e at 3.35 TB/s of aggregate memory bandwidth. This architecture eliminates inter-GPU communication bottlenecks, enabling near-linear scaling across transformer model training tasks that demand rapid gradient synchronization and massive parameter counts.

The XD675 is designed for datacenter-scale deployment within the HPE Cray supercomputer ecosystem, integrating tightly with HPE's high-speed Slingshot 11 interconnect fabric to enable multi-node scaling across thousands of GPU endpoints. Dual fourth-generation AMD EPYC processors provide the CPU compute backbone, supplying ample PCIe Gen 5 lanes, high memory bandwidth, and the I/O throughput necessary to keep eight SXM5 GPUs fully saturated with training data. The server supports up to 6 TB of DDR5 system memory, and its NVMe storage subsystem ensures rapid checkpoint saving and dataset streaming without creating I/O-bound bottlenecks during training runs.

Direct liquid cooling (DLC) is a first-class design element in the XD675, with chassis-integrated cold plates covering both GPUs and CPUs to sustain peak thermal dissipation at datacenter scale without reliance on airflow-intensive hot-aisle containment alone. This approach enables higher rack densities and aligns with modern sustainable datacenter power and cooling strategies. The platform is managed through HPE iLO 6 with integrated Redfish API support, enabling programmatic lifecycle management, telemetry streaming, and integration with HPE GreenLake cloud management services for unified fleet observability across on-premises AI infrastructure.

Ideal for

  • Large language model pre-training and fine-tuning at scale, leveraging NVLink-connected H200 SXM5 GPU clusters spanning hundreds of nodes via HPE Slingshot fabric
  • Generative AI inference serving for latency-sensitive enterprise applications requiring massive GPU memory capacity to load multi-hundred-billion-parameter models in full precision
  • Multimodal AI model development combining vision transformers and language models that benefit from the 1.12 TB aggregate HBM3e memory pool across eight GPUs
  • Scientific simulation and computational fluid dynamics workloads in national laboratories and research institutions requiring sustained FP64 double-precision throughput
  • Drug discovery and molecular dynamics simulations where GPU-accelerated quantum chemistry and protein folding models demand extreme memory bandwidth and capacity
  • Sovereign AI and enterprise AI cloud deployments where organizations require on-premises GPU infrastructure with full datacenter integration and programmatic management via Redfish and HPE GreenLake

Technical specifications

ManufacturerHPE
Product LineHPE Cray XD Series
ModelHPE Cray XD675
Form Factor4U, OCP Data Center Ready (DCR) form factor
GPU ConfigurationUp to 8x NVIDIA H200 SXM5 Tensor Core GPUs
GPU InterconnectNVLink 5th Generation, 900 GB/s bidirectional per GPU pair
GPU Memory (Total)Up to 1.12 TB HBM3e (8x 141 GB per GPU)
GPU Memory Bandwidth (Aggregate)Up to 3.35 TB/s aggregate across all 8 GPUs
CPUDual 4th Gen AMD EPYC processors (Genoa)
System MemoryUp to 6 TB DDR5, 12x DIMM slots per processor (24 slots total)
PCIe GenerationPCIe Gen 5
Network FabricHPE Slingshot 11 (200 Gb/s per port), supports multi-rail configurations
Host Fabric InterfaceUp to 4x HPE Slingshot NIC (200 Gb/s) for GPU-direct RDMA
Internal StorageUp to 4x NVMe M.2 SSD (boot) + front-accessible NVMe U.2 drive bays
Cooling TechnologyDirect Liquid Cooling (DLC) with integrated GPU and CPU cold plates; supplementary airflow fans
Power SupplyRedundant high-efficiency (Titanium) PSUs; supports high-density rack PDU integration
ManagementHPE iLO 6 with Redfish API, IPMI 2.0, HPE GreenLake cloud integration
Operating System SupportRed Hat Enterprise Linux, SUSE Linux Enterprise Server, Ubuntu, CentOS Stream
AI Software StackCompatible with NVIDIA AI Enterprise, CUDA 12.x, cuDNN, NCCL, PyTorch, TensorFlow, JAX
SecurityHPE Silicon Root of Trust, Secure Boot, TPM 2.0, encrypted storage support
Chassis Dimensions4U OCP DCR chassis; rack-optimized for high-density GPU pod configurations
Launch2025 Q1

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandHPE
CategoryServers
SKUHPE-R9G65A
Part NumberR9G65A
ConditionNew
Product LineHPE Cray XD Series
ModelHPE Cray XD675
Form Factor4U, OCP Data Center Ready (DCR) form factor
GPU ConfigurationUp to 8x NVIDIA H200 SXM5 Tensor Core GPUs
GPU InterconnectNVLink 5th Generation, 900 GB/s bidirectional per GPU pair
GPU Memory (Total)Up to 1.12 TB HBM3e (8x 141 GB per GPU)
GPU Memory Bandwidth (Aggregate)Up to 3.35 TB/s aggregate across all 8 GPUs
CPUDual 4th Gen AMD EPYC processors (Genoa)
System MemoryUp to 6 TB DDR5, 12x DIMM slots per processor (24 slots total)
PCIe GenerationPCIe Gen 5
Network FabricHPE Slingshot 11 (200 Gb/s per port), supports multi-rail configurations
Host Fabric InterfaceUp to 4x HPE Slingshot NIC (200 Gb/s) for GPU-direct RDMA
Internal StorageUp to 4x NVMe M.2 SSD (boot) + front-accessible NVMe U.2 drive bays
Cooling TechnologyDirect Liquid Cooling (DLC) with integrated GPU and CPU cold plates; supplementary airflow fans
Power SupplyRedundant high-efficiency (Titanium) PSUs; supports high-density rack PDU integration
ManagementHPE iLO 6 with Redfish API, IPMI 2.0, HPE GreenLake cloud integration
Operating System SupportRed Hat Enterprise Linux, SUSE Linux Enterprise Server, Ubuntu, CentOS Stream
AI Software StackCompatible with NVIDIA AI Enterprise, CUDA 12.x, cuDNN, NCCL, PyTorch, TensorFlow, JAX
SecurityHPE Silicon Root of Trust, Secure Boot, TPM 2.0, encrypted storage support
Chassis Dimensions4U OCP DCR chassis; rack-optimized for high-density GPU pod configurations
Launch2025 Q1

Frequently Asked Questions about HPE Cray XD675 AI Accelerator Server

What does the HPE Cray XD675 AI Accelerator Server do?

The HPE Cray XD675 AI Accelerator Server is built for enterprise data-center workloads — virtualization (VMware, Proxmox, Nutanix), private cloud, database hosting, and AI/ML training. It fits standard EIA-310 server racks and supports redundant PSUs and hot-swap drives common in production environments.

What are the headline specs of the HPE Cray XD675 AI Accelerator Server?

Key specifications for the HPE Cray XD675 AI Accelerator Server: new condition; manufacturer HPE; product line HPE Cray XD Series; model HPE Cray XD675; form factor 4U, OCP Data Center Ready (DCR) form factor; gpu configuration Up to 8x NVIDIA H200 SXM5 Tensor Core GPUs; gpu interconnect NVLink 5th Generation, 900 GB/s bidirectional per GPU pair. Manufacturer part number R9G65A. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.