Brand: Intel | Category: GPUs
SKU: QCT-G3-8OAM | Part #: QCT-G3-8OAM | MPN: QCT-G3-8OAM
Contact for Pricing — Request a Quote
The Quanta Computer Intel Gaudi 3 8-OAM HGX-Style Server (QCT-G3-8OAM) is a purpose-built AI training and inference platform integrating eight Intel Gaudi 3 accelerator modules in an OAM (Open Accelerator Module) form factor within an industry-standard HGX-compatible server chassis. Each Gaudi 3 accelerator is built on a 5nm process node and delivers substantial improvements in compute density and memory bandwidth over its predecessor, featuring 64 Tensor Processor Cores (TPCs) and Matrix Multiplication Engines (MMEs) optimized for mixed-precision deep learning operations including BF16, FP8, and FP32. The eight accelerators are interconnected via Intel's high-bandwidth, low-latency scale-up fabric using 24 ports of 200Gbps Ethernet per OAM, enabling direct all-to-all communication without requiring external switch infrastructure for within-node collective operations.
The QCT-G3-8OAM platform is engineered to address the demands of large-scale generative AI model training, large language model (LLM) fine-tuning, and high-throughput inference serving. Each Gaudi 3 OAM carries 128GB of HBM2e memory, yielding an aggregate of 1TB of HBM capacity across the full eight-accelerator configuration, alongside aggregate memory bandwidth well suited to the memory-bound kernels common in transformer-based workloads. The server integrates dual high-core-count Intel Xeon Scalable processors to handle host-side data preprocessing, orchestration, and I/O, with PCIe Gen 5 connectivity between host CPUs and the accelerator complex.
Designed for enterprise datacenter deployment, the QCT-G3 platform supports standard data center power and cooling infrastructure, with the OAM modules utilizing direct liquid cooling or high-airflow air cooling depending on deployment configuration. The system is compatible with Intel's Gaudi software ecosystem, including the Intel Gaudi Software Suite, Optimum Habana integration for Hugging Face workloads, and support for PyTorch and TensorFlow via SynapseAI. Omnixon Global supplies this platform to enterprise and hyperscale customers across the UAE, GCC, EMEA, and APAC regions.
| Manufacturer | Intel |
| Manufacturer Part Number | QCT-G3-8OAM |
| Accelerator Model | Intel Gaudi 3 |
| Number of Accelerators | 8 x OAM modules |
| Accelerator Process Node | 5nm |
| Tensor Processor Cores (TPC) per OAM | 64 |
| HBM Capacity per OAM | 128 GB HBM2e |
| Total Aggregate HBM Capacity | 1 TB (8 x 128 GB) |
| Scale-Up Interconnect per OAM | 24 x 200 Gbps Ethernet ports (built-in) |
| Scale-Out Network Ports per OAM | 2 x 400 Gbps Ethernet OSFP ports |
| Host CPU | Dual Intel Xeon Scalable Processors (5th Gen) |
| Host-to-Accelerator Interface | PCIe Gen 5 |
| Supported Precisions | FP32, BF16, FP16, FP8 |
| Chassis Form Factor | HGX-style OAM server |
| Software Ecosystem | Intel Gaudi SynapseAI SDK, PyTorch, TensorFlow, Optimum Habana |
| Framework Support | PyTorch, TensorFlow, Hugging Face Transformers (via Optimum Habana) |
| Cooling Support | Air cooling / Direct Liquid Cooling (DLC) depending on configuration |
| Target Workloads | AI training, LLM fine-tuning, generative AI inference, HPC |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Intel |
| Category | GPUs |
| SKU | QCT-G3-8OAM |
| Part Number | QCT-G3-8OAM |
| Condition | New |
| Manufacturer Part Number | QCT-G3-8OAM |
| Accelerator Model | Intel Gaudi 3 |
| Number of Accelerators | 8 x OAM modules |
| Accelerator Process Node | 5nm |
| Tensor Processor Cores (TPC) per OAM | 64 |
| HBM Capacity per OAM | 128 GB HBM2e |
| Total Aggregate HBM Capacity | 1 TB (8 x 128 GB) |
| Scale-Up Interconnect per OAM | 24 x 200 Gbps Ethernet ports (built-in) |
| Scale-Out Network Ports per OAM | 2 x 400 Gbps Ethernet OSFP ports |
| Host CPU | Dual Intel Xeon Scalable Processors (5th Gen) |
| Host-to-Accelerator Interface | PCIe Gen 5 |
| Supported Precisions | FP32, BF16, FP16, FP8 |
| Chassis Form Factor | HGX-style OAM server |
| Software Ecosystem | Intel Gaudi SynapseAI SDK, PyTorch, TensorFlow, Optimum Habana |
| Framework Support | PyTorch, TensorFlow, Hugging Face Transformers (via Optimum Habana) |
| Cooling Support | Air cooling / Direct Liquid Cooling (DLC) depending on configuration |
| Target Workloads | AI training, LLM fine-tuning, generative AI inference, HPC |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.