Brand: Gigabyte | Category: GPUs
SKU: 900-2G133-0040-000 | Part #: 900-2G133-0040-000 | MPN: 900-2G133-0040-000
Contact for Pricing — Request a Quote
The NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator is built on the NVIDIA Ada Lovelace architecture, delivering a substantial leap in compute density and energy efficiency for modern enterprise and data center environments. With 142 second-generation RT Cores, 568 fourth-generation Tensor Cores, and 18,176 CUDA cores, the L40S is engineered to handle the most demanding AI inference, training, and graphics workloads simultaneously within a single-slot, full-height, full-length form factor. The GPU connects via PCIe Gen 4 x16 and supports NVLink bridge configurations for multi-GPU scaling, making it highly adaptable to a wide range of server chassis and rack configurations.
At the heart of the L40S is 48GB of GDDR6 memory operating on a 384-bit bus with a memory bandwidth of 864 GB/s, providing the headroom required for large language models, high-resolution rendering pipelines, and complex simulation workloads. The card delivers up to 91.6 TFLOPS of FP32 performance and up to 362 TOPS of INT8 inference throughput with sparsity, while the Transformer Engine with FP8 precision enables accelerated training and inference of modern generative AI models. The L40S operates within a 350W TDP envelope, with power supplied through dual 8-pin PCIe connectors, and supports passive cooling for compatibility with airflow-managed data center environments.
Designed explicitly for enterprise data center deployment, the L40S differentiates itself from consumer-grade accelerators through its support for ECC memory, professional-grade reliability features, and NVIDIA's suite of enterprise software stacks including CUDA, TensorRT, and Omniverse. The Gigabyte-branded L40S (MPN: 900-2G133-0040-000) targets organizations in sectors including financial services, healthcare, media and entertainment, telecommunications, and cloud service provision, where sustained, mixed-workload GPU utilization at scale is a core operational requirement. Omnixon Global makes this accelerator available to enterprise buyers across the UAE, GCC, EMEA, and APAC regions.
| Manufacturer | Gigabyte |
| Manufacturer Part Number | 900-2G133-0040-000 |
| GPU Architecture | NVIDIA Ada Lovelace |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 (4th Generation) |
| RT Cores | 142 (2nd Generation) |
| Memory Capacity | 48 GB GDDR6 |
| Memory Bus Width | 384-bit |
| Memory Bandwidth | 864 GB/s |
| FP32 Performance | 91.6 TFLOPS |
| TF32 Tensor Core Performance | 183 TFLOPS (362 TFLOPS with sparsity) |
| INT8 Tensor Core Performance | 362 TOPS (724 TOPS with sparsity) |
| FP8 Tensor Core Performance | 724 TFLOPS (1457 TFLOPS with sparsity) |
| ECC Memory Support | Yes |
| TDP | 350 W |
| Power Connectors | 2x 8-pin PCIe |
| Host Interface | PCIe Gen 4 x16 |
| Form Factor | Full-Height, Full-Length (FHFL), Dual-Slot |
| Cooling | Passive (data center airflow) |
| Display Outputs | None |
| NVLink Support | Yes (NVLink Bridge, 2-GPU) |
| GPU Memory Error Correction | ECC |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Gigabyte |
| Category | GPUs |
| SKU | 900-2G133-0040-000 |
| Part Number | 900-2G133-0040-000 |
| Condition | New |
| Manufacturer Part Number | 900-2G133-0040-000 |
| GPU Architecture | NVIDIA Ada Lovelace |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 (4th Generation) |
| RT Cores | 142 (2nd Generation) |
| Memory Capacity | 48 GB GDDR6 |
| Memory Bus Width | 384-bit |
| Memory Bandwidth | 864 GB/s |
| FP32 Performance | 91.6 TFLOPS |
| TF32 Tensor Core Performance | 183 TFLOPS (362 TFLOPS with sparsity) |
| INT8 Tensor Core Performance | 362 TOPS (724 TOPS with sparsity) |
| FP8 Tensor Core Performance | 724 TFLOPS (1457 TFLOPS with sparsity) |
| ECC Memory Support | Yes |
| TDP | 350 W |
| Power Connectors | 2x 8-pin PCIe |
| Host Interface | PCIe Gen 4 x16 |
| Form Factor | Full-Height, Full-Length (FHFL), Dual-Slot |
| Cooling | Passive (data center airflow) |
| Display Outputs | None |
| NVLink Support | Yes (NVLink Bridge, 2-GPU) |
| GPU Memory Error Correction | ECC |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.