Brand: NVIDIA | Category: GPUs
SKU: NVIDIA-900-2G133-0000-000 | Part #: 900-2G133-0000-000 | MPN: 900-2G133-0000-000
Contact for Pricing — Request a Quote
The NVIDIA L40S PCIe GPU (part number 900-2G133-0000-000) is a full-height single-slot, passively cooled accelerator built on the NVIDIA Ada Lovelace architecture for enterprise AI and high-performance computing environments. With 48 GB of GDDR6 memory and 864 GB/s memory bandwidth, the L40S delivers 1,457 TFLOPS of peak tensor performance in TF32 precision, making it suitable for demanding workloads that require both capacity and throughput without active cooling overhead. The GPU connects via PCIe 4.0 x192 and draws up to 350 W, powered by dual 6-pin 12V server-class connectors that integrate seamlessly into standard enterprise infrastructure.
NVIDIA's L40S is the preferred choice for AI infrastructure teams deploying large language model training, inference pipelines, and mixed HPC workloads at scale. The passive air-cooled design eliminates blower noise and maintenance concerns, while the single-slot form factor maximizes GPU density in multi-GPU server configurations. Every NVIDIA L40S (900-2G133-0000-000) includes a three-year manufacturer warranty, with extended coverage options available. To source this accelerator for your enterprise deployment, contact Omnixon Global for an RFQ.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | NVIDIA-900-2G133-0000-000 |
| Part Number | 900-2G133-0000-000 |
| Condition | New |
| Capacity | 48GB |
| Interface | PCIe |
| GPU Model | L40S |
| Power Connector | 12V-2x6 / 16-pin (server-class) |
| Warranty | 3-year manufacturer warranty (extended available) |
| Manufacturer Part Number | 900-2G133-0000-000 |
| Product Line | NVIDIA L40S |
| Form Factor | PCIe Full-Height Single-Slot GPU |
| GPU Memory | 48 GB GDDR6 |
| Memory Bandwidth | 864 GB/s |
| Peak Tensor Performance (TF32) | 1,457 TFLOPS |
| Interface | PCIe 4.0 x192 |
| Maximum Power Consumption | 350 W |
| Architecture | NVIDIA Ada Lovelace |
| Primary Use Cases | LLM Training, Inference, HPC Workloads |
| Cooling | Passive (air-cooled) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote authorised-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold authorised channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.