HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator

HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator

Brand: HPE | Category: GPUs

SKU: P65893-B21 | Part #: P65893-B21 | MPN: P65893-B21

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator

The HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator (P65893-B21) is a high-performance AI training and inference accelerator built on Intel's third-generation Gaudi architecture. Designed for Open Accelerator Module (OAM) form factor deployments, the card integrates 96GB of HBM2e memory across a wide memory bus, delivering the high bandwidth and capacity essential for large-scale deep learning model training, generative AI workloads, and complex inference serving at scale. The Gaudi 3 architecture features an array of Tensor Processor Cores (TPCs) and Matrix Multiplication Engines (MMEs) purpose-built to accelerate transformer-based and convolutional neural network workloads with strong compute efficiency.

Gaudi 3 introduces significant improvements over its predecessor in both compute throughput and interconnect capability. The accelerator supports onboard RoCE v2-based RDMA networking via integrated 200Gb Ethernet ports, enabling direct accelerator-to-accelerator communication across nodes without relying on a separate networking fabric. This tightly coupled networking approach reduces communication bottlenecks during distributed training runs, making it well suited for multi-node clusters scaling to hundreds of accelerators. The OAM form factor enables dense system configurations within HPE's supported server and accelerator tray platforms.

As an HPE-branded solution, P65893-B21 is integrated into HPE's enterprise server and management ecosystem, supporting deployment alongside HPE iLO management, HPE GreenLake cloud services, and HPE's AI-optimized infrastructure portfolio. The accelerator is targeted at organizations building on-premises AI infrastructure for large language model (LLM) training, fine-tuning, and inference, as well as high-performance computing environments where scalable, standards-based AI acceleration is required. Intel's open software ecosystem, including the Intel Gaudi Software suite and SynapseAI framework, provides broad compatibility with PyTorch and other leading ML frameworks.

Ideal for

  • Large language model (LLM) pre-training and fine-tuning at scale across multi-node HPE accelerator clusters
  • Generative AI inference serving for enterprise applications requiring high-throughput, low-latency model execution
  • Computer vision and multimodal AI model training for healthcare imaging, manufacturing inspection, and media analytics
  • High-performance computing workloads combining AI and simulation in research, energy, and scientific computing environments
  • Retrieval-augmented generation (RAG) pipeline acceleration for enterprise knowledge management and intelligent search
  • Distributed deep learning research and development on on-premises AI infrastructure with open, standards-based tooling

Technical specifications

ManufacturerHPE
Manufacturer Part NumberP65893-B21
Accelerator ArchitectureIntel Gaudi 3
Form FactorOAM (Open Accelerator Module)
HBM Capacity96 GB
Memory TypeHBM2e
Memory Bandwidth3.7 TB/s
BF16 Compute Throughput1835 TFLOPS
FP8 Compute Throughput3670 TFLOPS
Integrated Network Ports24x 100GbE (aggregated to 21x 200GbE RDMA-capable ports)
Network ProtocolRoCE v2 (RDMA over Converged Ethernet)
Interconnect Bandwidth2.1 Tb/s bidirectional (per accelerator)
TDP (Thermal Design Power)900W
Supported FrameworksPyTorch, TensorFlow (via Intel SynapseAI / Gaudi Software)
Management IntegrationHPE iLO compatible server ecosystem
Process NodeTSMC 5nm
Supported Precision FormatsFP8, BF16, FP16, FP32, TF32
Platform CompatibilityHPE OAM-compatible accelerator tray and server platforms

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandHPE
CategoryGPUs
SKUP65893-B21
Part NumberP65893-B21
ConditionNew
Manufacturer Part NumberP65893-B21
Accelerator ArchitectureIntel Gaudi 3
Form FactorOAM (Open Accelerator Module)
HBM Capacity96 GB
Memory TypeHBM2e
Memory Bandwidth3.7 TB/s
BF16 Compute Throughput1835 TFLOPS
FP8 Compute Throughput3670 TFLOPS
Integrated Network Ports24x 100GbE (aggregated to 21x 200GbE RDMA-capable ports)
Network ProtocolRoCE v2 (RDMA over Converged Ethernet)
Interconnect Bandwidth2.1 Tb/s bidirectional (per accelerator)
TDP (Thermal Design Power)900W
Supported FrameworksPyTorch, TensorFlow (via Intel SynapseAI / Gaudi Software)
Management IntegrationHPE iLO compatible server ecosystem
Process NodeTSMC 5nm
Supported Precision FormatsFP8, BF16, FP16, FP32, TF32
Platform CompatibilityHPE OAM-compatible accelerator tray and server platforms

Frequently Asked Questions about HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator

What server platforms accept the HPE Intel Gaudi 3 OAM 96GB HBM2e Accelerator?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.