Brand: Dell | Category: GPUs
SKU: L4-PCIE-24GB | Part #: L4-PCIE-24GB | MPN: L4-PCIE-24GB
Contact for Pricing — Request a Quote
The NVIDIA L4 24GB GDDR6 PCIe 4.0, configured for Dell PowerEdge R760 and R650 platforms, is built on NVIDIA's Ada Lovelace architecture and delivers exceptional versatility across AI inference, video transcoding, virtual workstation, and data analytics workloads. With 7,680 CUDA cores, 240 fourth-generation Tensor cores, and 58 third-generation RT cores, the L4 achieves a remarkable balance of compute density and power efficiency within a single-slot, 75W passive form factor — making it one of the most deployment-friendly accelerators available for enterprise rack environments.
The L4's 24GB GDDR6 memory pool, operating at 300 GB/s bandwidth, is large enough to hold substantial AI models entirely on-device, enabling low-latency inference for large language model serving, computer vision pipelines, and recommendation engines. NVIDIA's fourth-generation NVENC and third-generation NVDEC engines allow the card to handle multiple simultaneous 4K and 8K video streams for media and entertainment or streaming analytics use cases. The card's PCIe 4.0 x16 interface ensures it integrates seamlessly into the high-throughput I/O architecture of the Dell PowerEdge R760 and R650 server platforms without requiring additional power connectors.
Designed explicitly for hyperscale and enterprise datacenter deployment, the L4 supports NVIDIA's full software ecosystem including CUDA 12.x, TensorRT, Triton Inference Server, and the NVIDIA AI Enterprise software suite. Its passive cooling design is optimized for the forced-air thermal environments of standard 1U and 2U rack servers, and multiple L4 cards can be installed within a single host system where slot and power budget allow, enabling flexible scale-out inference infrastructure without the space or power overhead of larger accelerators.
| Manufacturer | Dell |
| Manufacturer Part Number | L4-PCIE-24GB |
| GPU Architecture | NVIDIA Ada Lovelace |
| CUDA Cores | 7,680 |
| Tensor Cores | 240 (4th Generation) |
| RT Cores | 58 (3rd Generation) |
| GPU Memory | 24 GB GDDR6 |
| Memory Bandwidth | 300 GB/s |
| Memory Interface | 192-bit |
| Interface | PCIe 4.0 x16 |
| Form Factor | Single-slot, full-height half-length (FHHL), passive cooling |
| Thermal Design Power (TDP) | 72W |
| Power Connector | None (slot-powered) |
| FP32 Peak Performance | 30.3 TFLOPS |
| TF32 Tensor Core Performance | 120 TFLOPS (sparsity: 242 TFLOPS) |
| FP16 Tensor Core Performance | 242 TFLOPS (sparsity: 485 TFLOPS) |
| INT8 Tensor Core Performance | 485 TOPS (sparsity: 970 TOPS) |
| Video Encode Engines | 2x 4th Generation NVENC |
| Video Decode Engines | 1x 3rd Generation NVDEC |
| Compatible Server Platforms | Dell PowerEdge R760, Dell PowerEdge R650 |
| Operating System Support | Windows Server, Red Hat Enterprise Linux, Ubuntu Server, VMware vSphere (with NVIDIA vGPU software) |
| NVIDIA Software Ecosystem | CUDA 12.x, TensorRT, Triton Inference Server, NVIDIA AI Enterprise |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Dell |
| Category | GPUs |
| SKU | L4-PCIE-24GB |
| Part Number | L4-PCIE-24GB |
| Condition | New |
| Manufacturer Part Number | L4-PCIE-24GB |
| GPU Architecture | NVIDIA Ada Lovelace |
| CUDA Cores | 7,680 |
| Tensor Cores | 240 (4th Generation) |
| RT Cores | 58 (3rd Generation) |
| GPU Memory | 24 GB GDDR6 |
| Memory Bandwidth | 300 GB/s |
| Memory Interface | 192-bit |
| Interface | PCIe 4.0 x16 |
| Form Factor | Single-slot, full-height half-length (FHHL), passive cooling |
| Thermal Design Power (TDP) | 72W |
| Power Connector | None (slot-powered) |
| FP32 Peak Performance | 30.3 TFLOPS |
| TF32 Tensor Core Performance | 120 TFLOPS (sparsity: 242 TFLOPS) |
| FP16 Tensor Core Performance | 242 TFLOPS (sparsity: 485 TFLOPS) |
| INT8 Tensor Core Performance | 485 TOPS (sparsity: 970 TOPS) |
| Video Encode Engines | 2x 4th Generation NVENC |
| Video Decode Engines | 1x 3rd Generation NVDEC |
| Compatible Server Platforms | Dell PowerEdge R760, Dell PowerEdge R650 |
| Operating System Support | Windows Server, Red Hat Enterprise Linux, Ubuntu Server, VMware vSphere (with NVIDIA vGPU software) |
| NVIDIA Software Ecosystem | CUDA 12.x, TensorRT, Triton Inference Server, NVIDIA AI Enterprise |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.