Brand: Intel | Category: GPUs
SKU: P65396-B21 | Part #: P65396-B21 | MPN: P65396-B21
Contact for Pricing — Request a Quote
The HPE ProLiant DL380 Gen11 Intel Gaudi 3 Accelerator Option (P65396-B21) integrates Intel's Gaudi 3 AI accelerator architecture directly into the HPE ProLiant DL380 Gen11 platform, delivering purpose-built deep learning training and inference performance for enterprise data center environments. Gaudi 3 is built on a 5nm process node and features 64 Tensor Processor Cores alongside 8 matrix multiplication engines, with 128 GB of high-bandwidth memory (HBM2e) per accelerator card providing the memory capacity required for large-scale model training and serving. The architecture incorporates 24 integrated 200 GbE RDMA-capable network ports per accelerator, enabling high-throughput, low-latency scale-out communication without reliance on a separate networking fabric component.
This option kit is engineered specifically for compatibility and validated operation within the HPE ProLiant DL380 Gen11 server ecosystem, ensuring full integration with HPE's system management tooling, including iLO 6 and HPE OneView, for unified monitoring and lifecycle management of accelerator resources alongside host compute infrastructure. The Gaudi 3 accelerator supports the Intel Gaudi software stack, which provides compatibility with PyTorch and TensorFlow frameworks through the SynapseAI SDK, allowing enterprises to migrate existing AI workloads without significant code refactoring.
Targeted at organizations running large language model (LLM) training, generative AI inference, and high-performance deep learning pipelines, the DL380 Gen11 Intel Gaudi 3 Accelerator Option positions enterprise IT teams to deploy scalable AI infrastructure using an open, standards-aligned software ecosystem. Its fit within the widely deployed DL380 Gen11 2U form factor makes it operationally consistent with existing data center rack layouts, power delivery infrastructure, and HPE support structures across UAE, GCC, EMEA, and APAC deployments.
| Manufacturer | Intel |
| HPE Part Number | P65396-B21 |
| Product Name | HPE ProLiant DL380 Gen11 Intel Gaudi 3 Accelerator Option |
| Accelerator Architecture | Intel Gaudi 3 |
| Process Node | 5nm |
| Tensor Processor Cores | 64 |
| Matrix Multiplication Engines | 8 |
| High-Bandwidth Memory Type | HBM2e |
| Total HBM Capacity | 128 GB |
| Integrated Network Ports | 24x 200 GbE RDMA (per accelerator) |
| Network Protocol Support | RoCE v2 (RDMA over Converged Ethernet) |
| Software Framework Support | PyTorch, TensorFlow via Intel SynapseAI SDK |
| Host Server Compatibility | HPE ProLiant DL380 Gen11 |
| Form Factor | PCIe Accelerator Option Kit (2U server-compatible) |
| Management Integration | HPE iLO 6, HPE OneView |
| Interface | PCIe Gen5 |
| Operating System Support | Linux (distributions supported via Intel Gaudi software stack) |
| Target Workloads | AI training, LLM inference, deep learning, generative AI |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Intel |
| Category | GPUs |
| SKU | P65396-B21 |
| Part Number | P65396-B21 |
| Condition | New |
| HPE Part Number | P65396-B21 |
| Product Name | HPE ProLiant DL380 Gen11 Intel Gaudi 3 Accelerator Option |
| Accelerator Architecture | Intel Gaudi 3 |
| Process Node | 5nm |
| Tensor Processor Cores | 64 |
| Matrix Multiplication Engines | 8 |
| High-Bandwidth Memory Type | HBM2e |
| Total HBM Capacity | 128 GB |
| Integrated Network Ports | 24x 200 GbE RDMA (per accelerator) |
| Network Protocol Support | RoCE v2 (RDMA over Converged Ethernet) |
| Software Framework Support | PyTorch, TensorFlow via Intel SynapseAI SDK |
| Host Server Compatibility | HPE ProLiant DL380 Gen11 |
| Form Factor | PCIe Accelerator Option Kit (2U server-compatible) |
| Management Integration | HPE iLO 6, HPE OneView |
| Interface | PCIe Gen5 |
| Operating System Support | Linux (distributions supported via Intel Gaudi software stack) |
| Target Workloads | AI training, LLM inference, deep learning, generative AI |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.