Brand: NVIDIA | Category: GPUs
SKU: TCSL40S-PB | Part #: TCSL40S-PB | MPN: TCSL40S-PB
Contact for Pricing — Request a Quote
The NVIDIA L40S is a full-height, dual-slot GPU built on the Ada architecture, delivering 18,176 CUDA cores and 568 Tensor cores optimized for mixed-precision workloads. It provides 48GB of GDDR6 memory with 960 GB/s bandwidth, enabling efficient processing of large language models, multimodal AI applications, and graphics rendering at scale.
Designed as a universal data-center accelerator, the L40S excels across generative AI training and inference pipelines, supporting both FP8 and FP32 precisions alongside structured sparsity. The architecture balances compute throughput with memory capacity, making it suitable for enterprises deploying transformers, diffusion models, and real-time inference services without requiring model quantization or sharding across multiple GPUs.
The L40S integrates NVIDIA NVLink technology (two NVLinks per GPU for up to 900 GB/s peer-to-peer bandwidth) and supports virtualization via NVIDIA vGPU, enabling efficient multi-tenant cloud environments and consolidated workload scheduling across heterogeneous AI and visualization jobs.
| Manufacturer | NVIDIA |
| Brand | PNY |
| Model | L40S |
| ManufacturerPartNumber | TCSL40S-PB |
| GPU Memory | 48 GB GDDR6 |
| Memory Bandwidth | 960 GB/s |
| Memory Interface | 384-bit |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 |
| Boost Clock | 2.505 GHz |
| Max Power Consumption | 350W |
| Architecture | NVIDIA Ada |
| NVLink Connections | 2x NVLink (900 GB/s peer-to-peer) |
| Form Factor | Full Height, Dual Slot |
| PCIe Generation | PCIe 4.0 x16 |
| Compute Capability | 8.9 |
| Max Concurrent Threads Per Block | 1,024 |
| Launch Quarter | 2023-Q4 |
| Virtualization Support | NVIDIA vGPU |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | TCSL40S-PB |
| Part Number | TCSL40S-PB |
| Condition | New |
| Model | L40S |
| ManufacturerPartNumber | TCSL40S-PB |
| GPU Memory | 48 GB GDDR6 |
| Memory Bandwidth | 960 GB/s |
| Memory Interface | 384-bit |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 |
| Boost Clock | 2.505 GHz |
| Max Power Consumption | 350W |
| Architecture | NVIDIA Ada |
| NVLink Connections | 2x NVLink (900 GB/s peer-to-peer) |
| Form Factor | Full Height, Dual Slot |
| PCIe Generation | PCIe 4.0 x16 |
| Compute Capability | 8.9 |
| Max Concurrent Threads Per Block | 1,024 |
| Launch Quarter | 2023-Q4 |
| Virtualization Support | NVIDIA vGPU |