Brand: Intel | Category: GPUs
SKU: HL-325H | Part #: HL-325H | MPN: HL-325H
Contact for Pricing — Request a Quote
The Intel Gaudi 3 HL-325H is an OAM (OCP Accelerator Module) form-factor AI accelerator built on Intel's third-generation Gaudi architecture, fabricated on a 5nm process node. The HL-325H integrates 128GB of HBM2e memory across eight HBM2e stacks, delivering 3.7TB/s of aggregate HBM memory bandwidth. The accelerator features 64 Tensor Processor Cores (TPCs) and a dedicated Matrix Multiplication Engine (MME), providing substantial throughput for both training and inference workloads across FP8, BF16, FP16, and FP32 precisions. With 1835 TOPS of FP8 peak compute, the HL-325H is positioned for demanding large-scale AI model development in high-density datacenter deployments.
The Gaudi 3 architecture incorporates 24 integrated 100GbE RoCE v2 network ports directly on-die, enabling scale-out communication between accelerator nodes without requiring third-party networking silicon. This architecture supports direct server-to-server connectivity at up to 2.4Tb/s of total bisectional network bandwidth per accelerator, reducing fabric latency and infrastructure complexity in large AI clusters. The OAM mechanical form factor conforms to OCP OAM specifications, facilitating integration into OAM-compliant Universal Baseboard (UBB) systems from multiple platform vendors.
The HL-325H is supported by Intel's Gaudi software ecosystem, including the Habana SynapseAI SDK, which provides integration with industry-standard frameworks including PyTorch and TensorFlow. The accelerator is suited for enterprises deploying foundation model training, large language model fine-tuning, and high-throughput inference serving at scale across datacenter and cloud infrastructure environments in regulated and performance-sensitive industries.
| Manufacturer | Intel |
| Product Line | Gaudi 3 |
| Model | HL-325H |
| Form Factor | OCP OAM (OCP Accelerator Module) |
| Process Node | 5nm |
| Tensor Processor Cores (TPCs) | 64 |
| Peak FP8 Compute | 1835 TOPS |
| Peak BF16 Compute | 1835 TFLOPS |
| HBM Capacity | 128 GB |
| HBM Type | HBM2e |
| HBM Memory Bandwidth | 3.7 TB/s |
| On-Die Network Ports | 24x 100GbE RoCE v2 |
| Total Accelerator Network Bandwidth | 2.4 Tb/s |
| Supported Precisions | FP8, BF16, FP16, FP32 |
| Host Interface | PCIe Gen5 x16 |
| TDP | 900W |
| Software SDK | Intel Habana SynapseAI SDK |
| Framework Support | PyTorch, TensorFlow |
| OAM Specification Compliance | OCP OAM v1.0 |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Intel |
| Category | GPUs |
| SKU | HL-325H |
| Part Number | HL-325H |
| Condition | New |
| Product Line | Gaudi 3 |
| Model | HL-325H |
| Form Factor | OCP OAM (OCP Accelerator Module) |
| Process Node | 5nm |
| Tensor Processor Cores (TPCs) | 64 |
| Peak FP8 Compute | 1835 TOPS |
| Peak BF16 Compute | 1835 TFLOPS |
| HBM Capacity | 128 GB |
| HBM Type | HBM2e |
| HBM Memory Bandwidth | 3.7 TB/s |
| On-Die Network Ports | 24x 100GbE RoCE v2 |
| Total Accelerator Network Bandwidth | 2.4 Tb/s |
| Supported Precisions | FP8, BF16, FP16, FP32 |
| Host Interface | PCIe Gen5 x16 |
| TDP | 900W |
| Software SDK | Intel Habana SynapseAI SDK |
| Framework Support | PyTorch, TensorFlow |
| OAM Specification Compliance | OCP OAM v1.0 |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.