Gigabyte G262-IR1-AAX1 2-GPU L40S PCIe Server SKU

Gigabyte G262-IR1-AAX1 2-GPU L40S PCIe Server SKU

Brand: Gigabyte | Category: GPUs

SKU: G262-IR1-AAX1 | Part #: G262-IR1-AAX1 | MPN: G262-IR1-AAX1

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Gigabyte G262-IR1-AAX1 2-GPU L40S PCIe Server SKU

The Gigabyte G262-IR1-AAX1 is a 2U rack-mount server engineered specifically to support dual NVIDIA L40S PCIe graphics accelerators, delivering a high-density GPU compute platform optimized for enterprise AI inference, professional visualization, and data-intensive workloads. Built on Gigabyte's proven G262 server chassis lineage, this SKU pairs the computational density of two full-length PCIe GPU slots with a robust server architecture capable of sustaining the sustained thermal and power demands of the NVIDIA L40S — a GPU built on the Ada Lovelace architecture with third-generation RT Cores, fourth-generation Tensor Cores, and 48 GB of GDDR6 ECC memory per card.

The G262-IR1-AAX1 is designed to serve as a versatile multi-function GPU server, supporting NVIDIA's Ada Lovelace-generation L40S accelerators which deliver up to 91.6 TFLOPS of FP32 performance and 733 TOPS of INT8 inference throughput per card. With two L40S GPUs installed, the platform aggregates 96 GB of total GPU memory and provides substantial AI inference and training bandwidth within a 2U footprint. The server's PCIe architecture enables straightforward integration into existing data center fabric without requiring NVLink-based topology, making it compatible with a broad range of enterprise HPC and AI deployment patterns.

Targeted at datacenter operators, enterprise AI teams, and cloud service providers across the UAE, GCC, EMEA, and APAC regions, the G262-IR1-AAX1 supports workloads spanning large language model (LLM) inference, generative AI serving, 3D rendering pipelines, video transcoding at scale, and virtual workstation hosting. The server's 2U form factor optimizes rack space utilization while preserving sufficient airflow headroom for the L40S GPUs' 350 W TDP per card, and the platform supports standard data center power and management infrastructure including IPMI-based remote management.

Ideal for

  • Large language model (LLM) inference serving using both L40S GPUs in a high-throughput, low-latency production environment
  • Generative AI application hosting, including image synthesis and multi-modal model inference, leveraging Ada Lovelace Tensor Core performance
  • Professional 3D visualization and GPU-accelerated rendering pipelines for media, engineering, and architecture workflows
  • Virtual workstation and VDI deployments using NVIDIA vGPU software on L40S to support multiple concurrent professional users
  • AI-accelerated video transcoding and streaming analytics at scale within broadcast, media, and surveillance infrastructure
  • Enterprise HPC workloads including scientific simulation and data analytics requiring large aggregate GPU memory across PCIe-attached accelerators

Technical specifications

ManufacturerGigabyte
Manufacturer Part NumberG262-IR1-AAX1
Form Factor2U Rack-Mount Server
GPU Configuration2 x NVIDIA L40S PCIe (dual-slot)
GPU ArchitectureNVIDIA Ada Lovelace
GPU Memory per Card48 GB GDDR6 ECC
Total Aggregate GPU Memory96 GB GDDR6 ECC (2 GPUs)
FP32 Performance per L40S91.6 TFLOPS
INT8 Inference Throughput per L40S733 TOPS
GPU InterconnectPCIe Gen 4
GPU TDP per Card350 W
RT Cores Generation3rd Generation (Ada Lovelace)
Tensor Core Generation4th Generation (Ada Lovelace)
Server Chipset PlatformIntel Xeon Scalable (4th Gen, Sapphire Rapids) — G262-IR1 platform
CPU SocketDual LGA4677
PCIe SlotsMultiple PCIe Gen 5 / Gen 4 slots
Storage InterfaceNVMe and SATA support via onboard controllers
NetworkOnboard dual-port Ethernet (management and data); supports OCP 3.0 NIC expansion
Power SupplyRedundant hot-swap PSUs
Remote ManagementIPMI 2.0 / Gigabyte MegaRAC SP-X (BMC)
Operating System SupportLinux (major enterprise distributions); Windows Server
Rack Unit Height2U

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandGigabyte
CategoryGPUs
SKUG262-IR1-AAX1
Part NumberG262-IR1-AAX1
ConditionNew
Manufacturer Part NumberG262-IR1-AAX1
Form Factor2U Rack-Mount Server
GPU Configuration2 x NVIDIA L40S PCIe (dual-slot)
GPU ArchitectureNVIDIA Ada Lovelace
GPU Memory per Card48 GB GDDR6 ECC
Total Aggregate GPU Memory96 GB GDDR6 ECC (2 GPUs)
FP32 Performance per L40S91.6 TFLOPS
INT8 Inference Throughput per L40S733 TOPS
GPU InterconnectPCIe Gen 4
GPU TDP per Card350 W
RT Cores Generation3rd Generation (Ada Lovelace)
Tensor Core Generation4th Generation (Ada Lovelace)
Server Chipset PlatformIntel Xeon Scalable (4th Gen, Sapphire Rapids) — G262-IR1 platform
CPU SocketDual LGA4677
PCIe SlotsMultiple PCIe Gen 5 / Gen 4 slots
Storage InterfaceNVMe and SATA support via onboard controllers
NetworkOnboard dual-port Ethernet (management and data); supports OCP 3.0 NIC expansion
Power SupplyRedundant hot-swap PSUs
Remote ManagementIPMI 2.0 / Gigabyte MegaRAC SP-X (BMC)
Operating System SupportLinux (major enterprise distributions); Windows Server
Rack Unit Height2U

Frequently Asked Questions about Gigabyte G262-IR1-AAX1 2-GPU L40S PCIe Server SKU

What server platforms accept the Gigabyte G262-IR1-AAX1 2-GPU L40S PCIe Server SKU?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.