Intel Gaudi 2 HL-225B PCIe Accelerator Card 96GB HBM2e

Intel Gaudi 2 HL-225B PCIe Accelerator Card 96GB HBM2e

Brand: Intel | Category: GPUs

SKU: HL-225B | Part #: HL-225B | MPN: HL-225B

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Intel Gaudi 2 HL-225B PCIe Accelerator Card 96GB HBM2e

The Intel Gaudi 2 HL-225B is a PCIe Gen 4.0 accelerator card built on Intel's second-generation Gaudi architecture, designed to deliver high-throughput deep learning training and inference performance for enterprise datacenters. Featuring 96GB of HBM2e memory across six stacks with 2.45 TB/s of aggregate memory bandwidth, the HL-225B provides the memory capacity and bandwidth required to accommodate large language models, computer vision networks, and recommendation systems without model parallelism overhead. The card communicates over a standard PCIe x16 Gen 4.0 host interface, enabling broad server compatibility across major platforms without requiring proprietary interconnect infrastructure.

The Gaudi 2 architecture integrates 24 fully programmable Tensor Processor Cores (TPCs) alongside a Matrix Multiplication Engine (MME) optimized for GEMM operations central to transformer-based model training. Eight on-die 100 Gigabit Ethernet (RoCE v2) RDMA ports enable scale-out networking directly from the silicon, allowing multi-node training clusters to be constructed using standard Ethernet switching infrastructure rather than proprietary fabrics. The HL-225B variant routes these ports externally through an OSFP connector, supporting direct integration into high-bandwidth datacenter network topologies.

The card is supported by Intel's SynapseAI software suite, which provides a PyTorch and TensorFlow integration layer, a graph compiler, and runtime libraries that allow enterprises to migrate existing AI workloads with minimal code changes. The HL-225B is positioned for organizations seeking an alternative to dominant GPU incumbents for large-scale AI training infrastructure, particularly where open networking standards and total infrastructure flexibility are operational priorities. Omnixon Global supplies the HL-225B to enterprise buyers across the UAE, GCC, EMEA, and APAC regions.

Ideal for

  • Large language model (LLM) pre-training and fine-tuning at scale, leveraging 96GB HBM2e capacity to fit multi-billion parameter models on fewer accelerator cards
  • Enterprise AI inference serving for transformer-based models including BERT, GPT variants, and vision transformers requiring high memory bandwidth and low latency
  • Recommendation engine training for e-commerce, media, and financial services platforms where embedding table size demands substantial on-card HBM capacity
  • Computer vision model training for manufacturing quality inspection, medical imaging analysis, and autonomous systems using large batch sizes enabled by ample memory headroom
  • Multi-node distributed training clusters built over standard 100GbE RoCE v2 Ethernet fabric, eliminating dependency on proprietary high-speed interconnects
  • Datacenter AI infrastructure modernization projects requiring PCIe form-factor accelerators compatible with existing server chassis and power delivery systems

Technical specifications

ManufacturerIntel
Manufacturer Part NumberHL-225B
Product FamilyIntel Gaudi 2
ArchitectureGaudi 2
Form FactorPCIe Add-in Card (HHHL Full Height, Full Length)
Host InterfacePCIe Gen 4.0 x16
HBM Memory Capacity96 GB HBM2e
HBM Memory Stacks6 x HBM2e stacks
Memory Bandwidth2.45 TB/s
Tensor Processor Cores (TPC)24
On-Die Network Ports8 x 100 Gigabit Ethernet (RoCE v2)
External Networking ConnectorOSFP
Network ProtocolRDMA over Converged Ethernet v2 (RoCE v2)
Thermal Design Power (TDP)600 W
Power Connector2 x 8-pin PCIe auxiliary power
Supported FrameworksPyTorch, TensorFlow (via Intel SynapseAI SDK)
Operating System SupportLinux (Ubuntu, CentOS/RHEL)
Precision SupportFP32, BF16, FP16, INT16, INT8
Product SegmentEnterprise Datacenter AI Accelerator

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandIntel
CategoryGPUs
SKUHL-225B
Part NumberHL-225B
ConditionNew
Manufacturer Part NumberHL-225B
Product FamilyIntel Gaudi 2
ArchitectureGaudi 2
Form FactorPCIe Add-in Card (HHHL Full Height, Full Length)
Host InterfacePCIe Gen 4.0 x16
HBM Memory Capacity96 GB HBM2e
HBM Memory Stacks6 x HBM2e stacks
Memory Bandwidth2.45 TB/s
Tensor Processor Cores (TPC)24
On-Die Network Ports8 x 100 Gigabit Ethernet (RoCE v2)
External Networking ConnectorOSFP
Network ProtocolRDMA over Converged Ethernet v2 (RoCE v2)
Thermal Design Power (TDP)600 W
Power Connector2 x 8-pin PCIe auxiliary power
Supported FrameworksPyTorch, TensorFlow (via Intel SynapseAI SDK)
Operating System SupportLinux (Ubuntu, CentOS/RHEL)
Precision SupportFP32, BF16, FP16, INT16, INT8
Product SegmentEnterprise Datacenter AI Accelerator

Frequently Asked Questions about Intel Gaudi 2 HL-225B PCIe Accelerator Card 96GB HBM2e

What server platforms accept the Intel Gaudi 2 HL-225B PCIe Accelerator Card 96GB HBM2e?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.