Supermicro AS-8125GS-TNMR2 Intel Gaudi 2 Server

Supermicro AS-8125GS-TNMR2 Intel Gaudi 2 Server

Brand: Intel | Category: GPUs

SKU: SYS-821GS-TNMR2 | Part #: SYS-821GS-TNMR2 | MPN: SYS-821GS-TNMR2

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Supermicro AS-8125GS-TNMR2 Intel Gaudi 2 Server

The Supermicro AS-8125GS-TNMR2 (system model SYS-821GS-TNMR2) is a purpose-built AI training and inference server integrating Intel Gaudi 2 deep learning accelerators into a dense 8U rackmount chassis. The system is designed around Intel's Gaudi 2 HL-225H mezzanine compute cards, which are interconnected via a high-bandwidth on-board fabric using 21 x 100GbE RDMA-capable ports per accelerator — 24 internal ports for all-to-all server-level communication and 3 external ports for scale-out networking — eliminating the need for a separate high-speed switch fabric for many training cluster configurations.

Each Intel Gaudi 2 accelerator delivers substantial matrix multiplication throughput for mixed-precision AI workloads, supporting BF16 and FP32 natively, and is equipped with 96 GB of HBM2e memory per accelerator with up to 2.45 TB/s of aggregate memory bandwidth across the full complement of 8 accelerators in the system. The platform is validated for use with Intel's Habana SynapseAI software stack, providing support for popular deep learning frameworks including PyTorch and TensorFlow through a mature, production-grade software ecosystem.

Targeted at enterprise datacenters, cloud service providers, and AI research institutions, the AS-8125GS-TNMR2 supports dual 4th Gen Intel Xeon Scalable (Sapphire Rapids) processors to provide ample CPU-side compute for data preprocessing, orchestration, and host-side inference tasks. The system's architecture is optimized for large language model (LLM) training, generative AI model development, and high-throughput inference at scale, making it a competitive platform for organizations pursuing serious AI infrastructure buildouts across the UAE, GCC, EMEA, and APAC regions.

Ideal for

  • Large language model (LLM) pre-training and fine-tuning at scale using distributed data-parallel and model-parallel strategies across Gaudi 2 accelerator clusters
  • Generative AI application development and production inference serving for text, image, and multimodal foundation models
  • High-throughput deep learning research requiring dense HBM2e memory capacity for training transformer-based architectures with large batch sizes
  • Enterprise AI platform consolidation, enabling multiple simultaneous training jobs within a single high-density 8U chassis to maximize datacenter floor space efficiency
  • Scale-out AI cluster deployments leveraging Gaudi 2's native 100GbE RDMA fabric for low-latency, high-bandwidth inter-node communication without additional proprietary networking hardware
  • Regulated-industry AI workloads in financial services, healthcare, and government requiring on-premises accelerated compute with full data sovereignty

Technical specifications

ManufacturerIntel / Supermicro
System ModelAS-8125GS-TNMR2
Manufacturer Part NumberSYS-821GS-TNMR2
Form Factor8U Rackmount
AI AcceleratorIntel Gaudi 2 HL-225H
Number of Accelerators8
Accelerator Memory96 GB HBM2e per accelerator (768 GB total)
Accelerator Interconnect21 x 100GbE RDMA ports per Gaudi 2 (24 internal + 3 external scale-out)
Accelerator Precision SupportBF16, FP32, INT8
CPU SupportDual 4th Gen Intel Xeon Scalable Processors (Sapphire Rapids, Socket LGA4677)
Memory Slots32 x DDR5 DIMM slots
Maximum System Memory4 TB DDR5
Storage Bays8 x 2.5-inch NVMe/SATA hot-swap drive bays
Network (Host)2 x 10GbE BASE-T (onboard)
ManagementDedicated IPMI 2.0 / BMC management port
PCIe ExpansionPCIe 5.0 slots for additional host adapters
Power SupplyRedundant 3000W Titanium-level power supplies
Software StackIntel Habana SynapseAI (PyTorch, TensorFlow support)
Operating System SupportUbuntu 20.04 LTS, Ubuntu 22.04 LTS, Red Hat Enterprise Linux

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandIntel
CategoryGPUs
SKUSYS-821GS-TNMR2
Part NumberSYS-821GS-TNMR2
ConditionNew
System ModelAS-8125GS-TNMR2
Manufacturer Part NumberSYS-821GS-TNMR2
Form Factor8U Rackmount
AI AcceleratorIntel Gaudi 2 HL-225H
Number of Accelerators8
Accelerator Memory96 GB HBM2e per accelerator (768 GB total)
Accelerator Interconnect21 x 100GbE RDMA ports per Gaudi 2 (24 internal + 3 external scale-out)
Accelerator Precision SupportBF16, FP32, INT8
CPU SupportDual 4th Gen Intel Xeon Scalable Processors (Sapphire Rapids, Socket LGA4677)
Memory Slots32 x DDR5 DIMM slots
Maximum System Memory4 TB DDR5
Storage Bays8 x 2.5-inch NVMe/SATA hot-swap drive bays
Network (Host)2 x 10GbE BASE-T (onboard)
ManagementDedicated IPMI 2.0 / BMC management port
PCIe ExpansionPCIe 5.0 slots for additional host adapters
Power SupplyRedundant 3000W Titanium-level power supplies
Software StackIntel Habana SynapseAI (PyTorch, TensorFlow support)
Operating System SupportUbuntu 20.04 LTS, Ubuntu 22.04 LTS, Red Hat Enterprise Linux

Frequently Asked Questions about Supermicro AS-8125GS-TNMR2 Intel Gaudi 2 Server

What server platforms accept the Supermicro AS-8125GS-TNMR2 Intel Gaudi 2 Server?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.