NVIDIA L40S PCIe GPU 900-2G133-0000-000

NVIDIA L40S PCIe GPU 900-2G133-0000-000

Brand: NVIDIA | Category: GPUs

SKU: NVIDIA-900-2G133-0000-000 | Part #: 900-2G133-0000-000 | MPN: 900-2G133-0000-000

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the NVIDIA L40S PCIe GPU 900-2G133-0000-000

The NVIDIA L40S PCIe GPU (part number 900-2G133-0000-000) is a full-height single-slot, passively cooled accelerator built on the NVIDIA Ada Lovelace architecture for enterprise AI and high-performance computing environments. With 48 GB of GDDR6 memory and 864 GB/s memory bandwidth, the L40S delivers 1,457 TFLOPS of peak tensor performance in TF32 precision, making it suitable for demanding workloads that require both capacity and throughput without active cooling overhead. The GPU connects via PCIe 4.0 x192 and draws up to 350 W, powered by dual 6-pin 12V server-class connectors that integrate seamlessly into standard enterprise infrastructure.

NVIDIA's L40S is the preferred choice for AI infrastructure teams deploying large language model training, inference pipelines, and mixed HPC workloads at scale. The passive air-cooled design eliminates blower noise and maintenance concerns, while the single-slot form factor maximizes GPU density in multi-GPU server configurations. Every NVIDIA L40S (900-2G133-0000-000) includes a three-year manufacturer warranty, with extended coverage options available. To source this accelerator for your enterprise deployment, contact Omnixon Global for an RFQ.

Typical Workloads & Deployment Scenarios

  • Large Language Model (LLM) training and fine-tuning with high-capacity memory pools
  • Real-time inference serving for generative AI applications and transformer models
  • High-performance computing (HPC) simulations requiring sustained memory bandwidth
  • Multi-GPU cluster deployments leveraging single-slot passive cooling for thermal efficiency
  • Data center acceleration where noise reduction and space optimization are critical constraints

Technical Specifications

BrandNVIDIA
CategoryGPUs
SKUNVIDIA-900-2G133-0000-000
Part Number900-2G133-0000-000
ConditionNew
Capacity48GB
InterfacePCIe
GPU ModelL40S
Power Connector12V-2x6 / 16-pin (server-class)
Warranty3-year manufacturer warranty (extended available)
Manufacturer Part Number900-2G133-0000-000
Product LineNVIDIA L40S
Form FactorPCIe Full-Height Single-Slot GPU
GPU Memory48 GB GDDR6
Memory Bandwidth864 GB/s
Peak Tensor Performance (TF32)1,457 TFLOPS
InterfacePCIe 4.0 x192
Maximum Power Consumption350 W
ArchitectureNVIDIA Ada Lovelace
Primary Use CasesLLM Training, Inference, HPC Workloads
CoolingPassive (air-cooled)

Frequently Asked Questions about NVIDIA L40S PCIe GPU 900-2G133-0000-000

What server platforms accept the NVIDIA L40S PCIe GPU 900-2G133-0000-000?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote authorised-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold authorised channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.