Brand: Intel | Category: GPUs
SKU: P65395-B21 | Part #: P65395-B21 | MPN: P65395-B21
Contact for Pricing — Request a Quote
The HPE ProLiant DL380a Gen11 2x Gaudi 2 PCIe Server (P65395-B21) is a purpose-engineered 2U rack server designed to accelerate deep learning training and AI inference workloads at enterprise and hyperscale datacenter scale. The system integrates two Intel Gaudi 2 accelerator cards in a PCIe form factor, leveraging the Gaudi 2 HL-225H ASIC architecture with 96 GB of HBM2e memory per accelerator and 24 tensor processor cores per card, delivering high-throughput matrix computations suited to large language model training, computer vision, and generative AI pipelines.
Built on the HPE ProLiant DL380a Gen11 platform, the server pairs the dual Gaudi 2 accelerators with support for 4th Generation Intel Xeon Scalable processors and DDR5 system memory, providing the CPU-side computational bandwidth and memory capacity needed to feed AI accelerator workloads efficiently. The Gaudi 2 cards interconnect via on-card RDMA-capable 24-port 100 GbE RoCE v2 integration, enabling scale-out training across multiple nodes without requiring external InfiniBand fabric in many configurations. HPE iLO 6 management and the ProLiant Gen11 silicon root-of-trust security architecture are embedded throughout the platform.
Targeted at enterprise IT organizations, national AI research institutions, and datacenter operators across UAE, GCC, EMEA, and APAC, the DL380a Gen11 2x Gaudi 2 PCIe Server is optimized for organizations deploying the Intel Gaudi software ecosystem, including integration with PyTorch and TensorFlow via the Intel Gaudi software suite (SynapseAI), providing a vendor-supported software path for large-scale AI model development and production inference serving.
| Manufacturer | Intel (accelerator); Hewlett Packard Enterprise (server platform) |
| HPE Part Number | P65395-B21 |
| Server Form Factor | 2U Rack |
| Server Platform | HPE ProLiant DL380a Gen11 |
| Number of AI Accelerators | 2x Intel Gaudi 2 PCIe |
| Accelerator ASIC | Intel Gaudi 2 HL-225H |
| Accelerator Memory | 96 GB HBM2e per accelerator (192 GB total) |
| Tensor Processor Cores per Accelerator | 24 |
| Accelerator Interconnect | 24x 100 GbE RoCE v2 ports per card (on-card RDMA) |
| Accelerator Interface | PCIe Gen 4 |
| Supported CPU Generation | 4th Generation Intel Xeon Scalable Processors |
| System Memory Type | DDR5 |
| Management Controller | HPE iLO 6 |
| Security | HPE Silicon Root of Trust, measured boot, firmware integrity verification |
| AI Software Ecosystem | Intel SynapseAI SDK, PyTorch (Habana integration), TensorFlow |
| Operating System Support | Ubuntu, Red Hat Enterprise Linux (RHEL) |
| Network Fabric (Scale-Out) | RoCE v2 via on-card Gaudi 2 ports (no external InfiniBand required for node-level scale-out) |
| Chassis Standard | ANSI/EIA 310-D 19-inch rack compatible |
| Target Workloads | Deep learning training, LLM fine-tuning, generative AI inference, computer vision |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Intel |
| Category | GPUs |
| SKU | P65395-B21 |
| Part Number | P65395-B21 |
| Condition | New |
| HPE Part Number | P65395-B21 |
| Server Form Factor | 2U Rack |
| Server Platform | HPE ProLiant DL380a Gen11 |
| Number of AI Accelerators | 2x Intel Gaudi 2 PCIe |
| Accelerator ASIC | Intel Gaudi 2 HL-225H |
| Accelerator Memory | 96 GB HBM2e per accelerator (192 GB total) |
| Tensor Processor Cores per Accelerator | 24 |
| Accelerator Interconnect | 24x 100 GbE RoCE v2 ports per card (on-card RDMA) |
| Accelerator Interface | PCIe Gen 4 |
| Supported CPU Generation | 4th Generation Intel Xeon Scalable Processors |
| System Memory Type | DDR5 |
| Management Controller | HPE iLO 6 |
| Security | HPE Silicon Root of Trust, measured boot, firmware integrity verification |
| AI Software Ecosystem | Intel SynapseAI SDK, PyTorch (Habana integration), TensorFlow |
| Operating System Support | Ubuntu, Red Hat Enterprise Linux (RHEL) |
| Network Fabric (Scale-Out) | RoCE v2 via on-card Gaudi 2 ports (no external InfiniBand required for node-level scale-out) |
| Chassis Standard | ANSI/EIA 310-D 19-inch rack compatible |
| Target Workloads | Deep learning training, LLM fine-tuning, generative AI inference, computer vision |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.