Brand: Lenovo | Category: GPUs
SKU: LENO-4X67A84633 | Part #: 4X67A84633 | MPN: 4X67A84633
Contact for Pricing — Request a Quote
The Lenovo ThinkSystem GPU Module NVIDIA L40S 48GB PCIe brings the full capability of NVIDIA's Ada Lovelace architecture to validated ThinkSystem server platforms, delivering a purpose-built solution for organizations deploying large-scale generative AI inference, high-performance computing, and professional visualization workloads. Built on the third-generation Ada Lovelace GPU architecture, the L40S integrates 18,176 CUDA cores, 568 fourth-generation Tensor Cores, and 142 RT Cores within a 48GB GDDR6 frame buffer, yielding up to 91.6 TFLOPS of FP32 compute and 733 TOPS of INT8 inferencing throughput. The high-capacity onboard memory, paired with 864 GB/s memory bandwidth, ensures that large language models, diffusion models, and complex simulation workloads remain resident in GPU memory without costly data staging overhead.
Validated as a Lenovo ThinkSystem add-in card, the L40S is fully integrated into Lenovo's system firmware, management tooling, and thermal designs, ensuring predictable performance characteristics in rackmount server environments. The module supports PCIe Gen 4 x16 connectivity and is compatible with multi-GPU configurations via NVLink bridge (where platform topology allows), enabling scaling of AI inference capacity proportionally with deployment needs. Its 300W TDP is managed through active cooling solutions validated by Lenovo engineering, maintaining stability across sustained enterprise workloads.
Launched in Q2 2025, this ThinkSystem-validated module targets enterprises that require a single-slot, PCIe-native accelerator without the overhead of NVSwitch fabric infrastructure, making it particularly attractive for inference-optimized server clusters, virtual workstation infrastructure via NVIDIA vGPU software, and hybrid AI-plus-visualization deployments. The combination of Ada Lovelace Tensor Cores, dedicated AV1 hardware encode/decode engines, and NVIDIA's mature ecosystem of AI frameworks — including TensorRT, Triton Inference Server, and CUDA — positions the L40S as a versatile, production-grade accelerator within Lenovo's enterprise AI portfolio.
| Manufacturer | Lenovo |
| Product Line | ThinkSystem GPU Module |
| GPU Chip | NVIDIA L40S (Ada Lovelace Architecture) |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 (4th Generation) |
| RT Cores | 142 (3rd Generation) |
| GPU Memory | 48GB GDDR6 with ECC |
| Memory Bandwidth | 864 GB/s |
| FP32 Performance | 91.6 TFLOPS |
| FP16 Performance (Tensor) | 362.05 TFLOPS (with sparsity: 724.1 TFLOPS) |
| INT8 Throughput | 733 TOPS (with sparsity: 1457 TOPS) |
| FP64 Performance | 1.45 TFLOPS |
| Interface | PCIe Gen 4 x16 |
| Form Factor | Dual-slot, full-height full-length (FHFL) add-in card |
| Thermal Design Power (TDP) | 300W |
| Cooling | Active (server-validated forced-air cooling) |
| Display Outputs | 4x DisplayPort 1.4a |
| Multi-GPU Support | NVLink (platform-dependent topology) |
| vGPU Support | NVIDIA vGPU software compatible (NVIDIA AI Enterprise license required) |
| Video Encode/Decode | AV1, H.265, H.264 hardware encode and decode engines |
| Validated Platforms | Lenovo ThinkSystem Gen 3 and Gen 4 servers |
| Launch Generation | 2025-Q2 |
| OS / Framework Support | Linux, Windows Server; CUDA, TensorRT, Triton Inference Server, PyTorch, TensorFlow |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Lenovo |
| Category | GPUs |
| SKU | LENO-4X67A84633 |
| Part Number | 4X67A84633 |
| Condition | New |
| Product Line | ThinkSystem GPU Module |
| GPU Chip | NVIDIA L40S (Ada Lovelace Architecture) |
| CUDA Cores | 18,176 |
| Tensor Cores | 568 (4th Generation) |
| RT Cores | 142 (3rd Generation) |
| GPU Memory | 48GB GDDR6 with ECC |
| Memory Bandwidth | 864 GB/s |
| FP32 Performance | 91.6 TFLOPS |
| FP16 Performance (Tensor) | 362.05 TFLOPS (with sparsity: 724.1 TFLOPS) |
| INT8 Throughput | 733 TOPS (with sparsity: 1457 TOPS) |
| FP64 Performance | 1.45 TFLOPS |
| Interface | PCIe Gen 4 x16 |
| Form Factor | Dual-slot, full-height full-length (FHFL) add-in card |
| Thermal Design Power (TDP) | 300W |
| Cooling | Active (server-validated forced-air cooling) |
| Display Outputs | 4x DisplayPort 1.4a |
| Multi-GPU Support | NVLink (platform-dependent topology) |
| vGPU Support | NVIDIA vGPU software compatible (NVIDIA AI Enterprise license required) |
| Video Encode/Decode | AV1, H.265, H.264 hardware encode and decode engines |
| Validated Platforms | Lenovo ThinkSystem Gen 3 and Gen 4 servers |
| Launch Generation | 2025-Q2 |
| OS / Framework Support | Linux, Windows Server; CUDA, TensorRT, Triton Inference Server, PyTorch, TensorFlow |
The Lenovo ThinkSystem GPU Module NVIDIA L40S 48GB PCIe accelerates AI/ML training, inference, scientific HPC, and virtualization (vGPU) workloads. Typical deployments include LLM training clusters, computer-vision pipelines, financial risk modeling, and rendering farms.
Key specifications for the Lenovo ThinkSystem GPU Module NVIDIA L40S 48GB PCIe: new condition; manufacturer Lenovo; product line ThinkSystem GPU Module; gpu chip NVIDIA L40S (Ada Lovelace Architecture); cuda cores 18,176; tensor cores 568 (4th Generation); rt cores 142 (3rd Generation). Manufacturer part number 4X67A84633. For the full datasheet with electrical, environmental, and compliance details, contact our pre-sales engineering team.
The Lenovo ThinkSystem GPU Module NVIDIA L40S 48GB PCIe requires a PCIe Gen4 or Gen5 x16 slot, server power adequate for the card's TDP, and CUDA/ROCm driver support in your hypervisor or bare-metal OS. Sales engineering will confirm chassis fit (1U/2U/4U), PCIe lane count, and PSU headroom before quoting.