HPE ProLiant Compute DL384 Gen12

HPE ProLiant Compute DL384 Gen12

Brand: HPE | Category: Servers

SKU: HP-P75312B21 | Part #: P75312-B21 | MPN: P75312-B21

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the HPE ProLiant Compute DL384 Gen12

The HPE ProLiant Compute DL384 Gen12 represents a new class of AI-optimized infrastructure, engineered around the NVIDIA GH200 Grace Hopper Superchip architecture. Each GH200 module tightly integrates a 72-core ARM Neoverse V2 Grace CPU with an NVIDIA Hopper H100 GPU via a high-bandwidth NVLink-C2C interconnect delivering 900 GB/s of bidirectional bandwidth, effectively eliminating the traditional PCIe bottleneck that constrains GPU memory-access performance in conventional server designs. The DL384 Gen12 houses multiple GH200 Superchips in a dense, thermally engineered 4U chassis with direct liquid cooling support, enabling sustained full-performance operation across all accelerators simultaneously without thermal throttling.

Designed explicitly for generative AI, large language model (LLM) training, and high-throughput inference workloads, the DL384 Gen12 leverages the massive unified HBM3e memory pool available per GH200 module—up to 96 GB HBM3e per GPU combined with up to 480 GB LPDDR5X CPU memory per Superchip—providing an unprecedented addressable memory space that enables in-memory operation of multi-hundred-billion parameter models without model sharding across nodes. NVSwitch fabric interconnects across multiple Superchips within the system deliver NVLink-based all-to-all GPU communication, dramatically accelerating collective operations such as AllReduce that are critical in distributed training runs.

HPE integrates the DL384 Gen12 with its ProLiant ecosystem management stack including iLO 7 with AI-assisted telemetry, HPE Integrated Smart Update Manager, and native support for HPE GreenLake cloud services for hybrid AI infrastructure management. The server is validated for deployment within HPE Cray supercomputing clusters as well as standalone rack configurations, supporting NVIDIA Base Command Manager, NVIDIA AI Enterprise software suite, and major MLOps platforms. Its front-to-rear airflow and direct liquid cooling (DLC) architecture make it suitable for high-density AI data centers targeting Power Usage Effectiveness (PUE) efficiency at scale.

Ideal for

  • Training and fine-tuning large language models with 70B+ parameters entirely within unified GPU-CPU HBM3e memory without multi-node model parallelism overhead
  • High-throughput generative AI inference serving for enterprise LLM APIs, reducing token latency through near-zero PCIe memory transfer bottlenecks
  • Scientific simulation and high-performance computing workloads requiring tightly coupled CPU-GPU compute such as molecular dynamics, climate modeling, and computational fluid dynamics
  • Multimodal AI model development including vision-language models and diffusion-based image and video generation requiring large GPU memory footprints
  • Enterprise AI platform deployments running NVIDIA AI Enterprise on-premises with HPE GreenLake hybrid cloud management for secure, governed model development pipelines
  • Retrieval-augmented generation (RAG) pipeline acceleration where large vector databases and LLM inference must co-reside in memory for sub-millisecond retrieval-to-generation latency

Technical specifications

ManufacturerHP (Hewlett Packard Enterprise)
Product LineHPE ProLiant Compute
ModelDL384 Gen12
Form Factor4U Rack Server
Accelerator ArchitectureNVIDIA GH200 Grace Hopper Superchip
Number of GH200 SuperchipsUp to 8 x NVIDIA GH200 Superchips
GPU Compute per SuperchipNVIDIA H100 Hopper GPU (SXM5-class), 80 SM / 132 SM Hopper architecture
GPU Memory per Superchip96 GB HBM3e at 3.35 TB/s bandwidth
CPU per Superchip72-core NVIDIA Grace (ARM Neoverse V2) at up to 3.1 GHz
CPU Memory per SuperchipUp to 480 GB LPDDR5X at 512 GB/s
CPU-GPU InterconnectNVLink-C2C, 900 GB/s bidirectional per Superchip
GPU-to-GPU InterconnectNVLink 4.0 / NVSwitch 3rd Gen fabric, up to 900 GB/s per GPU all-to-all
Total System GPU Memory (8x GH200)Up to 768 GB HBM3e
Network I/OUp to 4 x NVIDIA ConnectX-7 NDR 400GbE / InfiniBand 400Gb/s ports
StorageUp to 8 x 2.5-inch NVMe SSD bays (PCIe Gen5), supports HPE SCM and U.3 drives
System ManagementHPE iLO 7 with AI-assisted telemetry, Redfish API, SNMP, HPE Integrated Smart Update Manager (ISSUM)
CoolingN+1 redundant hot-plug fans with direct liquid cooling (DLC) rear-door heat exchanger support; compliant with ASHRAE A3/A4 with DLC
Power SupplyRedundant 2+2 Titanium-efficiency (96%+) HPE Flex Slot PSUs, up to 3200W each
OS SupportRed Hat Enterprise Linux 9.x, Ubuntu 22.04/24.04 LTS, SUSE Linux Enterprise Server 15 SP5+, VMware vSphere (GPU passthrough)
Software EcosystemNVIDIA AI Enterprise 5.x, NVIDIA Base Command Manager, CUDA 12.x, cuDNN 9.x, TensorRT-LLM, HPE GreenLake for Compute Ops Management
Chassis Dimensions4U, 447mm (W) x 1752mm max depth x 175mm (H)
GenerationGen12 (launched 2025 Q3)

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandHPE
CategoryServers
SKUHP-P75312B21
Part NumberP75312-B21
ConditionNew
Product LineHPE ProLiant Compute
ModelDL384 Gen12
Form Factor4U Rack Server
Accelerator ArchitectureNVIDIA GH200 Grace Hopper Superchip
Number of GH200 SuperchipsUp to 8 x NVIDIA GH200 Superchips
GPU Compute per SuperchipNVIDIA H100 Hopper GPU (SXM5-class), 80 SM / 132 SM Hopper architecture
GPU Memory per Superchip96 GB HBM3e at 3.35 TB/s bandwidth
CPU per Superchip72-core NVIDIA Grace (ARM Neoverse V2) at up to 3.1 GHz
CPU Memory per SuperchipUp to 480 GB LPDDR5X at 512 GB/s
CPU-GPU InterconnectNVLink-C2C, 900 GB/s bidirectional per Superchip
GPU-to-GPU InterconnectNVLink 4.0 / NVSwitch 3rd Gen fabric, up to 900 GB/s per GPU all-to-all
Total System GPU Memory (8x GH200)Up to 768 GB HBM3e
Network I/OUp to 4 x NVIDIA ConnectX-7 NDR 400GbE / InfiniBand 400Gb/s ports
StorageUp to 8 x 2.5-inch NVMe SSD bays (PCIe Gen5), supports HPE SCM and U.3 drives
System ManagementHPE iLO 7 with AI-assisted telemetry, Redfish API, SNMP, HPE Integrated Smart Update Manager (ISSUM)
CoolingN+1 redundant hot-plug fans with direct liquid cooling (DLC) rear-door heat exchanger support; compliant with ASHRAE A3/A4 with DLC
Power SupplyRedundant 2+2 Titanium-efficiency (96%+) HPE Flex Slot PSUs, up to 3200W each
OS SupportRed Hat Enterprise Linux 9.x, Ubuntu 22.04/24.04 LTS, SUSE Linux Enterprise Server 15 SP5+, VMware vSphere (GPU passthrough)
Software EcosystemNVIDIA AI Enterprise 5.x, NVIDIA Base Command Manager, CUDA 12.x, cuDNN 9.x, TensorRT-LLM, HPE GreenLake for Compute Ops Management
Chassis Dimensions4U, 447mm (W) x 1752mm max depth x 175mm (H)
GenerationGen12 (launched 2025 Q3)

Frequently Asked Questions about HPE ProLiant Compute DL384 Gen12

What does the HPE ProLiant Compute DL384 Gen12 do?

The HPE ProLiant Compute DL384 Gen12 is built for enterprise data-center workloads — virtualization (VMware, Proxmox, Nutanix), private cloud, database hosting, and AI/ML training. It fits standard EIA-310 server racks and supports redundant PSUs and hot-swap drives common in production environments.

What are the headline specs of the HPE ProLiant Compute DL384 Gen12?

Key specifications for the HPE ProLiant Compute DL384 Gen12: new condition; manufacturer HP (Hewlett Packard Enterprise); product line HPE ProLiant Compute; model DL384 Gen12; form factor 4U Rack Server; accelerator architecture NVIDIA GH200 Grace Hopper Superchip; number of gh200 superchips Up to 8 x NVIDIA GH200 Superchips. Manufacturer part number P75312-B21. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.