Brand: NVIDIA | Category: GPUs
SKU: TCSL40S-PB | Part #: TCSL40S-PB | MPN: TCSL40S-PB
Contact for Pricing — Request a Quote
The NVIDIA L40S is a full-height, dual-slot GPU built on the Ada architecture, delivering 18,176 CUDA cores and 568 Tensor cores optimized for mixed-precision workloads. It provides 48GB of GDDR6 memory with 960 GB/s bandwidth, enabling efficient processing of large language models, multimodal AI applications, and graphics rendering at scale.
Designed as a universal data-center accelerator, the L40S excels across generative AI training and inference pipelines, supporting both FP8 and FP32 precisions alongside structured sparsity. The architecture balances compute throughput with memory capacity, making it suitable for enterprises deploying transformers, diffusion models, and real-time inference services without requiring model quantization or sharding across multiple GPUs.
The L40S integrates NVIDIA NVLink technology (two NVLinks per GPU for up to 900 GB/s peer-to-peer bandwidth) and supports virtualization via NVIDIA vGPU, enabling efficient multi-tenant cloud environments and consolidated workload scheduling across heterogeneous AI and visualization jobs.
| Manufacturer | NVIDIA |
| Brand | PNY |
| Model | L40S |
| ManufacturerPartNumber | TCSL40S-PB |
| GPU Memory | 48 GB GDDR6 |
| Memory Bandwidth | 960 GB/s |
| Memory Interface | 384-bit |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 |
| Boost Clock | 2.505 GHz |
| Max Power Consumption | 350W |
| Architecture | NVIDIA Ada |
| NVLink Connections | 2x NVLink (900 GB/s peer-to-peer) |
| Form Factor | Full Height, Dual Slot |
| PCIe Generation | PCIe 4.0 x16 |
| Compute Capability | 8.9 |
| Max Concurrent Threads Per Block | 1,024 |
| Launch Quarter | 2023-Q4 |
| Virtualization Support | NVIDIA vGPU |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | TCSL40S-PB |
| Part Number | TCSL40S-PB |
| Condition | New |
| Model | L40S |
| ManufacturerPartNumber | TCSL40S-PB |
| GPU Memory | 48 GB GDDR6 |
| Memory Bandwidth | 960 GB/s |
| Memory Interface | 384-bit |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 |
| Boost Clock | 2.505 GHz |
| Max Power Consumption | 350W |
| Architecture | NVIDIA Ada |
| NVLink Connections | 2x NVLink (900 GB/s peer-to-peer) |
| Form Factor | Full Height, Dual Slot |
| PCIe Generation | PCIe 4.0 x16 |
| Compute Capability | 8.9 |
| Max Concurrent Threads Per Block | 1,024 |
| Launch Quarter | 2023-Q4 |
| Virtualization Support | NVIDIA vGPU |
The PNY NVIDIA L40S 48GB accelerates AI/ML training, inference, scientific HPC, and virtualization (vGPU) workloads. Typical deployments include LLM training clusters, computer-vision pipelines, financial risk modeling, and rendering farms.
Key specifications for the PNY NVIDIA L40S 48GB: new condition; manufacturer NVIDIA; brand PNY; model L40S; manufacturerpartnumber TCSL40S-PB; gpu memory 48 GB GDDR6; memory bandwidth 960 GB/s. Manufacturer part number TCSL40S-PB. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.
The PNY NVIDIA L40S 48GB requires a PCIe Gen4 or Gen5 x16 slot, server power adequate for the card's TDP, and CUDA/ROCm driver support in your hypervisor or bare-metal OS. Sales engineering will confirm chassis fit (1U/2U/4U), PCIe lane count, and PSU headroom before quoting.