Brand: HPE | Category: GPUs
SKU: P65893-B21 | Part #: P65893-B21 | MPN: P65893-B21
Contact for Pricing — Request a Quote
The HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator (P65893-B21) is a high-performance AI training and inference accelerator built on Intel's third-generation Gaudi architecture. Designed for Open Accelerator Module (OAM) form factor deployments, the card integrates 96GB of HBM2e memory across a wide memory bus, delivering the high bandwidth and capacity essential for large-scale deep learning model training, generative AI workloads, and complex inference serving at scale. The Gaudi 3 architecture features an array of Tensor Processor Cores (TPCs) and Matrix Multiplication Engines (MMEs) purpose-built to accelerate transformer-based and convolutional neural network workloads with strong compute efficiency.
Gaudi 3 introduces significant improvements over its predecessor in both compute throughput and interconnect capability. The accelerator supports onboard RoCE v2-based RDMA networking via integrated 200Gb Ethernet ports, enabling direct accelerator-to-accelerator communication across nodes without relying on a separate networking fabric. This tightly coupled networking approach reduces communication bottlenecks during distributed training runs, making it well suited for multi-node clusters scaling to hundreds of accelerators. The OAM form factor enables dense system configurations within HPE's supported server and accelerator tray platforms.
As an HPE-branded solution, P65893-B21 is integrated into HPE's enterprise server and management ecosystem, supporting deployment alongside HPE iLO management, HPE GreenLake cloud services, and HPE's AI-optimized infrastructure portfolio. The accelerator is targeted at organizations building on-premises AI infrastructure for large language model (LLM) training, fine-tuning, and inference, as well as high-performance computing environments where scalable, standards-based AI acceleration is required. Intel's open software ecosystem, including the Intel Gaudi Software suite and SynapseAI framework, provides broad compatibility with PyTorch and other leading ML frameworks.
| Manufacturer | HPE |
| Manufacturer Part Number | P65893-B21 |
| Accelerator Architecture | Intel Gaudi 3 |
| Form Factor | OAM (Open Accelerator Module) |
| HBM Capacity | 96 GB |
| Memory Type | HBM2e |
| Memory Bandwidth | 3.7 TB/s |
| BF16 Compute Throughput | 1835 TFLOPS |
| FP8 Compute Throughput | 3670 TFLOPS |
| Integrated Network Ports | 24x 100GbE (aggregated to 21x 200GbE RDMA-capable ports) |
| Network Protocol | RoCE v2 (RDMA over Converged Ethernet) |
| Interconnect Bandwidth | 2.1 Tb/s bidirectional (per accelerator) |
| TDP (Thermal Design Power) | 900W |
| Supported Frameworks | PyTorch, TensorFlow (via Intel SynapseAI / Gaudi Software) |
| Management Integration | HPE iLO compatible server ecosystem |
| Process Node | TSMC 5nm |
| Supported Precision Formats | FP8, BF16, FP16, FP32, TF32 |
| Platform Compatibility | HPE OAM-compatible accelerator tray and server platforms |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P65893-B21 |
| Part Number | P65893-B21 |
| Condition | New |
| Manufacturer Part Number | P65893-B21 |
| Accelerator Architecture | Intel Gaudi 3 |
| Form Factor | OAM (Open Accelerator Module) |
| HBM Capacity | 96 GB |
| Memory Type | HBM2e |
| Memory Bandwidth | 3.7 TB/s |
| BF16 Compute Throughput | 1835 TFLOPS |
| FP8 Compute Throughput | 3670 TFLOPS |
| Integrated Network Ports | 24x 100GbE (aggregated to 21x 200GbE RDMA-capable ports) |
| Network Protocol | RoCE v2 (RDMA over Converged Ethernet) |
| Interconnect Bandwidth | 2.1 Tb/s bidirectional (per accelerator) |
| TDP (Thermal Design Power) | 900W |
| Supported Frameworks | PyTorch, TensorFlow (via Intel SynapseAI / Gaudi Software) |
| Management Integration | HPE iLO compatible server ecosystem |
| Process Node | TSMC 5nm |
| Supported Precision Formats | FP8, BF16, FP16, FP32, TF32 |
| Platform Compatibility | HPE OAM-compatible accelerator tray and server platforms |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.