Brand: Intel | Category: GPUs
SKU: 210-BKOF-G3PCIE | Part #: 210-BKOF-G3PCIE | MPN: 210-BKOF-G3PCIE
Contact for Pricing — Request a Quote
The Intel Gaudi 3 PCIe Inference Card delivers 96 GB of HBM2e memory in a PCIe Gen5 add-in card form factor, purpose-built for high-throughput AI inference and large language model serving. This Intel solution integrates 64 Tensor Processor Cores (TPCs) and 8 Matrix Multiplication Engines (MMEs) to accelerate deep learning workloads with exceptional efficiency. On-board networking includes 24 × 100 Gigabit Ethernet RDMA ports via integrated NICs, enabling seamless multi-card clustering and low-latency distributed inference across enterprise deployments. Active cooling via integrated blower fan ensures sustained performance in dense server environments.
AI infrastructure teams and data center architects responsible for LLM deployment pipelines increasingly standardize on Intel's Gaudi 3 architecture because the software stack natively supports PyTorch and Hugging Face frameworks while remaining fully compatible with vLLM serving stacks. The validated Dell PowerEdge R760xa configuration (part number 210-BKOF-G3PCIE) eliminates integration risk and accelerates time-to-production for organizations deploying inference at scale across Ubuntu and RHEL Linux environments. This PCIe card form factor allows customers to refresh inference capacity without wholesale server replacement, protecting existing capital investments across GCC and EMEA data centers.
Organizations planning multi-node inference clusters or seeking to consolidate LLM serving workloads should engage Omnixon Global's enterprise solutions team to discuss availability, volume pricing, and validated deployment architectures across UAE, GCC, EMEA, and APAC regions. Submit an RFQ today referencing Intel Gaudi 3 PCIe Inference Card 96GB (210-BKOF-G3PCIE) to receive configuration guidance and technical specifications tailored to your AI infrastructure roadmap.
```| Brand | Intel |
| Category | GPUs |
| SKU | 210-BKOF-G3PCIE |
| Part Number | 210-BKOF-G3PCIE |
| Condition | New |
| Manufacturer Part Number | 210-BKOF-G3PCIE |
| Product Line | Intel Gaudi 3 |
| Architecture | Gaudi 3 |
| Form Factor | PCIe Add-In Card |
| Host Interface | PCIe Gen5 |
| Memory Capacity | 96 GB |
| Memory Type | HBM2e |
| Number of Tensor Processor Cores (TPCs) | 64 |
| Number of Matrix Multiplication Engines (MMEs) | 8 |
| On-Card Networking | 24 × 100 Gigabit Ethernet RDMA ports (via integrated NICs) |
| Thermal Design | Active cooling (blower fan) |
| Software Stack | Intel Gaudi Software (PyTorch-native, supports Hugging Face, vLLM) |
| Target Workload | AI inference; large language model serving; deep learning |
| Server Compatibility | Dell PowerEdge R760xa (validated configuration) |
| Operating System Support | Linux (Ubuntu, RHEL) |
| Region Availability | UAE, GCC, EMEA, APAC |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.