PNY NVIDIA L40S 48GB

PNY NVIDIA L40S 48GB

Brand: NVIDIA | Category: GPUs

SKU: TCSL40S-PB | Part #: TCSL40S-PB | MPN: TCSL40S-PB

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the PNY NVIDIA L40S 48GB

The NVIDIA L40S is a full-height, dual-slot GPU built on the Ada architecture, delivering 18,176 CUDA cores and 568 Tensor cores optimized for mixed-precision workloads. It provides 48GB of GDDR6 memory with 960 GB/s bandwidth, enabling efficient processing of large language models, multimodal AI applications, and graphics rendering at scale.

Designed as a universal data-center accelerator, the L40S excels across generative AI training and inference pipelines, supporting both FP8 and FP32 precisions alongside structured sparsity. The architecture balances compute throughput with memory capacity, making it suitable for enterprises deploying transformers, diffusion models, and real-time inference services without requiring model quantization or sharding across multiple GPUs.

The L40S integrates NVIDIA NVLink technology (two NVLinks per GPU for up to 900 GB/s peer-to-peer bandwidth) and supports virtualization via NVIDIA vGPU, enabling efficient multi-tenant cloud environments and consolidated workload scheduling across heterogeneous AI and visualization jobs.

Ideal for

  • Large language model (LLM) fine-tuning and inference serving at 13B–70B parameter scales
  • Generative AI inference for text-to-image, image generation, and multimodal retrieval pipelines
  • Real-time ray tracing and complex 3D rendering for digital content creation and product visualization
  • Data-parallel and pipeline-parallel distributed training for custom transformers and foundation models
  • Multi-tenant GPU virtualization for consolidated enterprise AI and graphics workload management
  • Vector database embedding generation and semantic search at enterprise scale

Technical specifications

ManufacturerNVIDIA
BrandPNY
ModelL40S
ManufacturerPartNumberTCSL40S-PB
GPU Memory48 GB GDDR6
Memory Bandwidth960 GB/s
Memory Interface384-bit
CUDA Cores18,176
Tensor Cores568
Boost Clock2.505 GHz
Max Power Consumption350W
ArchitectureNVIDIA Ada
NVLink Connections2x NVLink (900 GB/s peer-to-peer)
Form FactorFull Height, Dual Slot
PCIe GenerationPCIe 4.0 x16
Compute Capability8.9
Max Concurrent Threads Per Block1,024
Launch Quarter2023-Q4
Virtualization SupportNVIDIA vGPU

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandNVIDIA
CategoryGPUs
SKUTCSL40S-PB
Part NumberTCSL40S-PB
ConditionNew
ModelL40S
ManufacturerPartNumberTCSL40S-PB
GPU Memory48 GB GDDR6
Memory Bandwidth960 GB/s
Memory Interface384-bit
CUDA Cores18,176
Tensor Cores568
Boost Clock2.505 GHz
Max Power Consumption350W
ArchitectureNVIDIA Ada
NVLink Connections2x NVLink (900 GB/s peer-to-peer)
Form FactorFull Height, Dual Slot
PCIe GenerationPCIe 4.0 x16
Compute Capability8.9
Max Concurrent Threads Per Block1,024
Launch Quarter2023-Q4
Virtualization SupportNVIDIA vGPU