Brand: Intel | Category: GPUs
SKU: SYS-821GS-TNMR2 | Part #: SYS-821GS-TNMR2 | MPN: SYS-821GS-TNMR2
Contact for Pricing — Request a Quote
The Supermicro AS-8125GS-TNMR2 (system model SYS-821GS-TNMR2) is a purpose-built AI training and inference server integrating Intel Gaudi 2 deep learning accelerators into a dense 8U rackmount chassis. The system is designed around Intel's Gaudi 2 HL-225H mezzanine compute cards, which are interconnected via a high-bandwidth on-board fabric using 21 x 100GbE RDMA-capable ports per accelerator — 24 internal ports for all-to-all server-level communication and 3 external ports for scale-out networking — eliminating the need for a separate high-speed switch fabric for many training cluster configurations.
Each Intel Gaudi 2 accelerator delivers substantial matrix multiplication throughput for mixed-precision AI workloads, supporting BF16 and FP32 natively, and is equipped with 96 GB of HBM2e memory per accelerator with up to 2.45 TB/s of aggregate memory bandwidth across the full complement of 8 accelerators in the system. The platform is validated for use with Intel's Habana SynapseAI software stack, providing support for popular deep learning frameworks including PyTorch and TensorFlow through a mature, production-grade software ecosystem.
Targeted at enterprise datacenters, cloud service providers, and AI research institutions, the AS-8125GS-TNMR2 supports dual 4th Gen Intel Xeon Scalable (Sapphire Rapids) processors to provide ample CPU-side compute for data preprocessing, orchestration, and host-side inference tasks. The system's architecture is optimized for large language model (LLM) training, generative AI model development, and high-throughput inference at scale, making it a competitive platform for organizations pursuing serious AI infrastructure buildouts across the UAE, GCC, EMEA, and APAC regions.
| Manufacturer | Intel / Supermicro |
| System Model | AS-8125GS-TNMR2 |
| Manufacturer Part Number | SYS-821GS-TNMR2 |
| Form Factor | 8U Rackmount |
| AI Accelerator | Intel Gaudi 2 HL-225H |
| Number of Accelerators | 8 |
| Accelerator Memory | 96 GB HBM2e per accelerator (768 GB total) |
| Accelerator Interconnect | 21 x 100GbE RDMA ports per Gaudi 2 (24 internal + 3 external scale-out) |
| Accelerator Precision Support | BF16, FP32, INT8 |
| CPU Support | Dual 4th Gen Intel Xeon Scalable Processors (Sapphire Rapids, Socket LGA4677) |
| Memory Slots | 32 x DDR5 DIMM slots |
| Maximum System Memory | 4 TB DDR5 |
| Storage Bays | 8 x 2.5-inch NVMe/SATA hot-swap drive bays |
| Network (Host) | 2 x 10GbE BASE-T (onboard) |
| Management | Dedicated IPMI 2.0 / BMC management port |
| PCIe Expansion | PCIe 5.0 slots for additional host adapters |
| Power Supply | Redundant 3000W Titanium-level power supplies |
| Software Stack | Intel Habana SynapseAI (PyTorch, TensorFlow support) |
| Operating System Support | Ubuntu 20.04 LTS, Ubuntu 22.04 LTS, Red Hat Enterprise Linux |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Intel |
| Category | GPUs |
| SKU | SYS-821GS-TNMR2 |
| Part Number | SYS-821GS-TNMR2 |
| Condition | New |
| System Model | AS-8125GS-TNMR2 |
| Manufacturer Part Number | SYS-821GS-TNMR2 |
| Form Factor | 8U Rackmount |
| AI Accelerator | Intel Gaudi 2 HL-225H |
| Number of Accelerators | 8 |
| Accelerator Memory | 96 GB HBM2e per accelerator (768 GB total) |
| Accelerator Interconnect | 21 x 100GbE RDMA ports per Gaudi 2 (24 internal + 3 external scale-out) |
| Accelerator Precision Support | BF16, FP32, INT8 |
| CPU Support | Dual 4th Gen Intel Xeon Scalable Processors (Sapphire Rapids, Socket LGA4677) |
| Memory Slots | 32 x DDR5 DIMM slots |
| Maximum System Memory | 4 TB DDR5 |
| Storage Bays | 8 x 2.5-inch NVMe/SATA hot-swap drive bays |
| Network (Host) | 2 x 10GbE BASE-T (onboard) |
| Management | Dedicated IPMI 2.0 / BMC management port |
| PCIe Expansion | PCIe 5.0 slots for additional host adapters |
| Power Supply | Redundant 3000W Titanium-level power supplies |
| Software Stack | Intel Habana SynapseAI (PyTorch, TensorFlow support) |
| Operating System Support | Ubuntu 20.04 LTS, Ubuntu 22.04 LTS, Red Hat Enterprise Linux |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.