NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator

NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator

Brand: Gigabyte | Category: GPUs

SKU: 900-2G133-0040-000 | Part #: 900-2G133-0040-000 | MPN: 900-2G133-0040-000

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator

The NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator is built on the NVIDIA Ada Lovelace architecture, delivering a substantial leap in compute density and energy efficiency for modern enterprise and data center environments. With 142 second-generation RT Cores, 568 fourth-generation Tensor Cores, and 18,176 CUDA cores, the L40S is engineered to handle the most demanding AI inference, training, and graphics workloads simultaneously within a single-slot, full-height, full-length form factor. The GPU connects via PCIe Gen 4 x16 and supports NVLink bridge configurations for multi-GPU scaling, making it highly adaptable to a wide range of server chassis and rack configurations.

At the heart of the L40S is 48GB of GDDR6 memory operating on a 384-bit bus with a memory bandwidth of 864 GB/s, providing the headroom required for large language models, high-resolution rendering pipelines, and complex simulation workloads. The card delivers up to 91.6 TFLOPS of FP32 performance and up to 362 TOPS of INT8 inference throughput with sparsity, while the Transformer Engine with FP8 precision enables accelerated training and inference of modern generative AI models. The L40S operates within a 350W TDP envelope, with power supplied through dual 8-pin PCIe connectors, and supports passive cooling for compatibility with airflow-managed data center environments.

Designed explicitly for enterprise data center deployment, the L40S differentiates itself from consumer-grade accelerators through its support for ECC memory, professional-grade reliability features, and NVIDIA's suite of enterprise software stacks including CUDA, TensorRT, and Omniverse. The Gigabyte-branded L40S (MPN: 900-2G133-0040-000) targets organizations in sectors including financial services, healthcare, media and entertainment, telecommunications, and cloud service provision, where sustained, mixed-workload GPU utilization at scale is a core operational requirement. Omnixon Global makes this accelerator available to enterprise buyers across the UAE, GCC, EMEA, and APAC regions.

Ideal for

  • Large-scale generative AI model training and fine-tuning, including large language models and diffusion-based image generation, leveraging FP8 Tensor Core throughput
  • High-throughput AI inference serving for real-time applications such as recommendation engines, natural language processing APIs, and computer vision pipelines
  • 3D visualization, ray tracing, and GPU-accelerated rendering for media production, architecture, and digital twin development using RT Core and CUDA resources
  • Scientific simulation and high-performance computing workloads in industries such as energy, life sciences, and computational fluid dynamics
  • Virtual workstation and cloud graphics delivery for demanding professional applications via GPU virtualization with NVIDIA vGPU software
  • Telecommunications and edge AI acceleration in centralized data center nodes supporting 5G network analytics, signal processing, and real-time inference

Technical specifications

ManufacturerGigabyte
Manufacturer Part Number900-2G133-0040-000
GPU ArchitectureNVIDIA Ada Lovelace
CUDA Cores18,176
Tensor Cores568 (4th Generation)
RT Cores142 (2nd Generation)
Memory Capacity48 GB GDDR6
Memory Bus Width384-bit
Memory Bandwidth864 GB/s
FP32 Performance91.6 TFLOPS
TF32 Tensor Core Performance183 TFLOPS (362 TFLOPS with sparsity)
INT8 Tensor Core Performance362 TOPS (724 TOPS with sparsity)
FP8 Tensor Core Performance724 TFLOPS (1457 TFLOPS with sparsity)
ECC Memory SupportYes
TDP350 W
Power Connectors2x 8-pin PCIe
Host InterfacePCIe Gen 4 x16
Form FactorFull-Height, Full-Length (FHFL), Dual-Slot
CoolingPassive (data center airflow)
Display OutputsNone
NVLink SupportYes (NVLink Bridge, 2-GPU)
GPU Memory Error CorrectionECC

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandGigabyte
CategoryGPUs
SKU900-2G133-0040-000
Part Number900-2G133-0040-000
ConditionNew
Manufacturer Part Number900-2G133-0040-000
GPU ArchitectureNVIDIA Ada Lovelace
CUDA Cores18,176
Tensor Cores568 (4th Generation)
RT Cores142 (2nd Generation)
Memory Capacity48 GB GDDR6
Memory Bus Width384-bit
Memory Bandwidth864 GB/s
FP32 Performance91.6 TFLOPS
TF32 Tensor Core Performance183 TFLOPS (362 TFLOPS with sparsity)
INT8 Tensor Core Performance362 TOPS (724 TOPS with sparsity)
FP8 Tensor Core Performance724 TFLOPS (1457 TFLOPS with sparsity)
ECC Memory SupportYes
TDP350 W
Power Connectors2x 8-pin PCIe
Host InterfacePCIe Gen 4 x16
Form FactorFull-Height, Full-Length (FHFL), Dual-Slot
CoolingPassive (data center airflow)
Display OutputsNone
NVLink SupportYes (NVLink Bridge, 2-GPU)
GPU Memory Error CorrectionECC

Frequently Asked Questions about NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator

What server platforms accept the NVIDIA L40S 48GB GDDR6 PCIe GPU Accelerator?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.