Omnixon Global
QCT QuantaGrid G3-8OAM Gaudi 3 Server System

QCT QuantaGrid G3-8OAM Gaudi 3 Server System

Brand: Intel | Category: GPUs

SKU: QCT-SG3-8OAM | Part #: QCT-SG3-8OAM | MPN: QCT-SG3-8OAM

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the QCT QuantaGrid G3-8OAM Gaudi 3 Server System

The QCT QuantaGrid G3-8OAM is a purpose-built AI accelerator server system designed around Intel's Gaudi 3 architecture, integrating eight Intel Gaudi 3 OAM (Open Accelerator Module) mezzanine accelerators into a high-density 5U chassis. The Gaudi 3 architecture delivers a significant generational leap in AI compute throughput, featuring 128 GB of HBM2e memory per OAM module, 64 matrix multiplication engines (MMEs), and 8 TPC (Tensor Processing Core) clusters per accelerator. The platform leverages Intel's 5 nm process node and provides high-bandwidth, low-latency scale-out connectivity through integrated 24-port 200 GbE RoCE networking per accelerator, enabling direct server-to-server communication without requiring a separate networking switch fabric for smaller clusters.

The G3-8OAM system is engineered for demanding generative AI training and large-scale inference workloads, delivering competitive FP8 and BF16 compute performance across the full OAM array. The server chassis accommodates dual Intel Xeon Scalable host CPUs, providing substantial system memory capacity and PCIe Gen 5 host connectivity to the accelerator subsystem. The OAM form factor and modular architecture align with OCP (Open Compute Project) standards, supporting ecosystem-wide tooling and operational consistency. The system ships with support for Intel Gaudi software stack, including the SynapseAI SDK, enabling compatibility with PyTorch and other major ML frameworks through a hardware abstraction layer.

Targeted at enterprise AI infrastructure, cloud service providers, and large-scale datacenter deployments across EMEA, GCC, UAE, and APAC regions, the QuantaGrid G3-8OAM is positioned as a full-system solution for organizations requiring scalable, standards-aligned AI compute. QCT's system integration brings together validated thermal management, power delivery rated for high sustained accelerator TDP, and datacenter-ready manageability interfaces including IPMI and Redfish-compliant BMC. Omnixon Global provides access to this platform for enterprise buyers across the UAE, GCC, EMEA, and APAC territories.

Ideal for

  • Large language model (LLM) pre-training and fine-tuning at scale, leveraging the combined HBM2e memory capacity and high-bandwidth inter-accelerator fabric across all eight Gaudi 3 OAM modules
  • Generative AI inference serving for production enterprise deployments requiring high-throughput FP8 and BF16 tensor operations with deterministic latency
  • Deep learning recommendation model (DLRM) training for e-commerce, media, and enterprise personalization platforms requiring dense embedding table operations
  • Computer vision and multimodal AI model training pipelines in healthcare imaging, autonomous systems, and industrial quality inspection
  • High-performance AI research and experimentation environments within enterprise datacenter or HPC clusters, utilizing RoCE-based scale-out networking for multi-node training jobs
  • Sovereign AI infrastructure buildouts for government, financial services, and regulated industries in UAE, GCC, and EMEA requiring locally operated, enterprise-grade AI compute

Technical specifications

ManufacturerIntel
Manufacturer Part NumberQCT-SG3-8OAM
System IntegratorQCT (Quanta Cloud Technology)
Product LineQuantaGrid
ModelG3-8OAM
AI AcceleratorIntel Gaudi 3 OAM
Number of Accelerators8x Intel Gaudi 3 OAM modules
Accelerator Memory128 GB HBM2e per OAM module (1 TB aggregate across 8 modules)
Accelerator Process NodeIntel 5 nm
Matrix Multiplication Engines (per accelerator)64 MMEs
Tensor Processing Cores (per accelerator)8 TPC clusters
Scale-Out Networking (per accelerator)24-port 200 GbE RoCE integrated per Gaudi 3 OAM
Host CPU SupportDual Intel Xeon Scalable (4th Gen or 5th Gen, socket dependent on configuration)
Host InterfacePCIe Gen 5
Form Factor5U rackmount
Accelerator Form FactorOAM (Open Accelerator Module) — OCP-aligned
Software StackIntel SynapseAI SDK; PyTorch support via Habana bridge
Management InterfaceIPMI, Redfish-compliant BMC
Target WorkloadsLLM training, generative AI inference, DLRM, computer vision, HPC AI
Regional AvailabilityUAE, GCC, EMEA, APAC

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandIntel
CategoryGPUs
SKUQCT-SG3-8OAM
Part NumberQCT-SG3-8OAM
ConditionNew
Manufacturer Part NumberQCT-SG3-8OAM
System IntegratorQCT (Quanta Cloud Technology)
Product LineQuantaGrid
ModelG3-8OAM
AI AcceleratorIntel Gaudi 3 OAM
Number of Accelerators8x Intel Gaudi 3 OAM modules
Accelerator Memory128 GB HBM2e per OAM module (1 TB aggregate across 8 modules)
Accelerator Process NodeIntel 5 nm
Matrix Multiplication Engines (per accelerator)64 MMEs
Tensor Processing Cores (per accelerator)8 TPC clusters
Scale-Out Networking (per accelerator)24-port 200 GbE RoCE integrated per Gaudi 3 OAM
Host CPU SupportDual Intel Xeon Scalable (4th Gen or 5th Gen, socket dependent on configuration)
Host InterfacePCIe Gen 5
Form Factor5U rackmount
Accelerator Form FactorOAM (Open Accelerator Module) — OCP-aligned
Software StackIntel SynapseAI SDK; PyTorch support via Habana bridge
Management InterfaceIPMI, Redfish-compliant BMC
Target WorkloadsLLM training, generative AI inference, DLRM, computer vision, HPC AI
Regional AvailabilityUAE, GCC, EMEA, APAC

Frequently Asked Questions about QCT QuantaGrid G3-8OAM Gaudi 3 Server System

What server platforms accept the QCT QuantaGrid G3-8OAM Gaudi 3 Server System?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.