Omnixon Global
Intel Gaudi 3 HL-325H OAM Accelerator 128GB HBM2e

Intel Gaudi 3 HL-325H OAM Accelerator 128GB HBM2e

Brand: Intel | Category: GPUs

SKU: HL-325H | Part #: HL-325H | MPN: HL-325H

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Intel Gaudi 3 HL-325H OAM Accelerator 128GB HBM2e

The Intel Gaudi 3 HL-325H is an OAM (OCP Accelerator Module) form-factor AI accelerator built on Intel's third-generation Gaudi architecture, fabricated on a 5nm process node. The HL-325H integrates 128GB of HBM2e memory across eight HBM2e stacks, delivering 3.7TB/s of aggregate HBM memory bandwidth. The accelerator features 64 Tensor Processor Cores (TPCs) and a dedicated Matrix Multiplication Engine (MME), providing substantial throughput for both training and inference workloads across FP8, BF16, FP16, and FP32 precisions. With 1835 TOPS of FP8 peak compute, the HL-325H is positioned for demanding large-scale AI model development in high-density datacenter deployments.

The Gaudi 3 architecture incorporates 24 integrated 100GbE RoCE v2 network ports directly on-die, enabling scale-out communication between accelerator nodes without requiring third-party networking silicon. This architecture supports direct server-to-server connectivity at up to 2.4Tb/s of total bisectional network bandwidth per accelerator, reducing fabric latency and infrastructure complexity in large AI clusters. The OAM mechanical form factor conforms to OCP OAM specifications, facilitating integration into OAM-compliant Universal Baseboard (UBB) systems from multiple platform vendors.

The HL-325H is supported by Intel's Gaudi software ecosystem, including the Habana SynapseAI SDK, which provides integration with industry-standard frameworks including PyTorch and TensorFlow. The accelerator is suited for enterprises deploying foundation model training, large language model fine-tuning, and high-throughput inference serving at scale across datacenter and cloud infrastructure environments in regulated and performance-sensitive industries.

Ideal for

  • Large language model (LLM) pre-training and fine-tuning at scale using multi-node Gaudi 3 clusters interconnected via on-chip RoCE v2 networking
  • High-throughput generative AI inference serving for enterprise applications requiring low-latency, high-concurrency response workloads
  • Computer vision and multimodal model training for industries including healthcare imaging, autonomous systems, and retail analytics
  • Recommendation system training and serving for large-scale e-commerce, media streaming, and personalization platforms
  • Scientific and research AI workloads in energy, genomics, and climate modeling that require high HBM bandwidth and FP32/BF16 compute precision
  • Enterprise MLOps pipelines requiring dense GPU-equivalent compute within OAM-compatible high-density baseboard server chassis

Technical specifications

ManufacturerIntel
Product LineGaudi 3
ModelHL-325H
Form FactorOCP OAM (OCP Accelerator Module)
Process Node5nm
Tensor Processor Cores (TPCs)64
Peak FP8 Compute1835 TOPS
Peak BF16 Compute1835 TFLOPS
HBM Capacity128 GB
HBM TypeHBM2e
HBM Memory Bandwidth3.7 TB/s
On-Die Network Ports24x 100GbE RoCE v2
Total Accelerator Network Bandwidth2.4 Tb/s
Supported PrecisionsFP8, BF16, FP16, FP32
Host InterfacePCIe Gen5 x16
TDP900W
Software SDKIntel Habana SynapseAI SDK
Framework SupportPyTorch, TensorFlow
OAM Specification ComplianceOCP OAM v1.0

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandIntel
CategoryGPUs
SKUHL-325H
Part NumberHL-325H
ConditionNew
Product LineGaudi 3
ModelHL-325H
Form FactorOCP OAM (OCP Accelerator Module)
Process Node5nm
Tensor Processor Cores (TPCs)64
Peak FP8 Compute1835 TOPS
Peak BF16 Compute1835 TFLOPS
HBM Capacity128 GB
HBM TypeHBM2e
HBM Memory Bandwidth3.7 TB/s
On-Die Network Ports24x 100GbE RoCE v2
Total Accelerator Network Bandwidth2.4 Tb/s
Supported PrecisionsFP8, BF16, FP16, FP32
Host InterfacePCIe Gen5 x16
TDP900W
Software SDKIntel Habana SynapseAI SDK
Framework SupportPyTorch, TensorFlow
OAM Specification ComplianceOCP OAM v1.0

Frequently Asked Questions about Intel Gaudi 3 HL-325H OAM Accelerator 128GB HBM2e

What server platforms accept the Intel Gaudi 3 HL-325H OAM Accelerator 128GB HBM2e?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.