Brand: Gigabyte | Category: GPUs
SKU: G262-IR1-AAX1 | Part #: G262-IR1-AAX1 | MPN: G262-IR1-AAX1
Contact for Pricing — Request a Quote
The Gigabyte G262-IR1-AAX1 is a 2U rack-mount server engineered specifically to support dual NVIDIA L40S PCIe graphics accelerators, delivering a high-density GPU compute platform optimized for enterprise AI inference, professional visualization, and data-intensive workloads. Built on Gigabyte's proven G262 server chassis lineage, this SKU pairs the computational density of two full-length PCIe GPU slots with a robust server architecture capable of sustaining the sustained thermal and power demands of the NVIDIA L40S — a GPU built on the Ada Lovelace architecture with third-generation RT Cores, fourth-generation Tensor Cores, and 48 GB of GDDR6 ECC memory per card.
The G262-IR1-AAX1 is designed to serve as a versatile multi-function GPU server, supporting NVIDIA's Ada Lovelace-generation L40S accelerators which deliver up to 91.6 TFLOPS of FP32 performance and 733 TOPS of INT8 inference throughput per card. With two L40S GPUs installed, the platform aggregates 96 GB of total GPU memory and provides substantial AI inference and training bandwidth within a 2U footprint. The server's PCIe architecture enables straightforward integration into existing data center fabric without requiring NVLink-based topology, making it compatible with a broad range of enterprise HPC and AI deployment patterns.
Targeted at datacenter operators, enterprise AI teams, and cloud service providers across the UAE, GCC, EMEA, and APAC regions, the G262-IR1-AAX1 supports workloads spanning large language model (LLM) inference, generative AI serving, 3D rendering pipelines, video transcoding at scale, and virtual workstation hosting. The server's 2U form factor optimizes rack space utilization while preserving sufficient airflow headroom for the L40S GPUs' 350 W TDP per card, and the platform supports standard data center power and management infrastructure including IPMI-based remote management.
| Manufacturer | Gigabyte |
| Manufacturer Part Number | G262-IR1-AAX1 |
| Form Factor | 2U Rack-Mount Server |
| GPU Configuration | 2 x NVIDIA L40S PCIe (dual-slot) |
| GPU Architecture | NVIDIA Ada Lovelace |
| GPU Memory per Card | 48 GB GDDR6 ECC |
| Total Aggregate GPU Memory | 96 GB GDDR6 ECC (2 GPUs) |
| FP32 Performance per L40S | 91.6 TFLOPS |
| INT8 Inference Throughput per L40S | 733 TOPS |
| GPU Interconnect | PCIe Gen 4 |
| GPU TDP per Card | 350 W |
| RT Cores Generation | 3rd Generation (Ada Lovelace) |
| Tensor Core Generation | 4th Generation (Ada Lovelace) |
| Server Chipset Platform | Intel Xeon Scalable (4th Gen, Sapphire Rapids) — G262-IR1 platform |
| CPU Socket | Dual LGA4677 |
| PCIe Slots | Multiple PCIe Gen 5 / Gen 4 slots |
| Storage Interface | NVMe and SATA support via onboard controllers |
| Network | Onboard dual-port Ethernet (management and data); supports OCP 3.0 NIC expansion |
| Power Supply | Redundant hot-swap PSUs |
| Remote Management | IPMI 2.0 / Gigabyte MegaRAC SP-X (BMC) |
| Operating System Support | Linux (major enterprise distributions); Windows Server |
| Rack Unit Height | 2U |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Gigabyte |
| Category | GPUs |
| SKU | G262-IR1-AAX1 |
| Part Number | G262-IR1-AAX1 |
| Condition | New |
| Manufacturer Part Number | G262-IR1-AAX1 |
| Form Factor | 2U Rack-Mount Server |
| GPU Configuration | 2 x NVIDIA L40S PCIe (dual-slot) |
| GPU Architecture | NVIDIA Ada Lovelace |
| GPU Memory per Card | 48 GB GDDR6 ECC |
| Total Aggregate GPU Memory | 96 GB GDDR6 ECC (2 GPUs) |
| FP32 Performance per L40S | 91.6 TFLOPS |
| INT8 Inference Throughput per L40S | 733 TOPS |
| GPU Interconnect | PCIe Gen 4 |
| GPU TDP per Card | 350 W |
| RT Cores Generation | 3rd Generation (Ada Lovelace) |
| Tensor Core Generation | 4th Generation (Ada Lovelace) |
| Server Chipset Platform | Intel Xeon Scalable (4th Gen, Sapphire Rapids) — G262-IR1 platform |
| CPU Socket | Dual LGA4677 |
| PCIe Slots | Multiple PCIe Gen 5 / Gen 4 slots |
| Storage Interface | NVMe and SATA support via onboard controllers |
| Network | Onboard dual-port Ethernet (management and data); supports OCP 3.0 NIC expansion |
| Power Supply | Redundant hot-swap PSUs |
| Remote Management | IPMI 2.0 / Gigabyte MegaRAC SP-X (BMC) |
| Operating System Support | Linux (major enterprise distributions); Windows Server |
| Rack Unit Height | 2U |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.