NVIDIA L4 24GB GDDR6 PCIe 4.0 – Dell PowerEdge R760 / R650

NVIDIA L4 24GB GDDR6 PCIe 4.0 – Dell PowerEdge R760 / R650

Brand: Dell | Category: GPUs

SKU: L4-PCIE-24GB | Part #: L4-PCIE-24GB | MPN: L4-PCIE-24GB

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the NVIDIA L4 24GB GDDR6 PCIe 4.0 – Dell PowerEdge R760 / R650

The NVIDIA L4 24GB GDDR6 PCIe 4.0, configured for Dell PowerEdge R760 and R650 platforms, is built on NVIDIA's Ada Lovelace architecture and delivers exceptional versatility across AI inference, video transcoding, virtual workstation, and data analytics workloads. With 7,680 CUDA cores, 240 fourth-generation Tensor cores, and 58 third-generation RT cores, the L4 achieves a remarkable balance of compute density and power efficiency within a single-slot, 75W passive form factor — making it one of the most deployment-friendly accelerators available for enterprise rack environments.

The L4's 24GB GDDR6 memory pool, operating at 300 GB/s bandwidth, is large enough to hold substantial AI models entirely on-device, enabling low-latency inference for large language model serving, computer vision pipelines, and recommendation engines. NVIDIA's fourth-generation NVENC and third-generation NVDEC engines allow the card to handle multiple simultaneous 4K and 8K video streams for media and entertainment or streaming analytics use cases. The card's PCIe 4.0 x16 interface ensures it integrates seamlessly into the high-throughput I/O architecture of the Dell PowerEdge R760 and R650 server platforms without requiring additional power connectors.

Designed explicitly for hyperscale and enterprise datacenter deployment, the L4 supports NVIDIA's full software ecosystem including CUDA 12.x, TensorRT, Triton Inference Server, and the NVIDIA AI Enterprise software suite. Its passive cooling design is optimized for the forced-air thermal environments of standard 1U and 2U rack servers, and multiple L4 cards can be installed within a single host system where slot and power budget allow, enabling flexible scale-out inference infrastructure without the space or power overhead of larger accelerators.

Ideal for

  • Large-scale AI inference serving for NLP, computer vision, and recommendation models in production datacenter environments
  • High-density video transcoding supporting multiple concurrent 4K and 8K streams for media, broadcast, and streaming analytics platforms
  • Virtual GPU (vGPU) deployment for GPU-accelerated virtual workstations and virtual desktop infrastructure across enterprise user populations
  • Real-time data analytics and machine learning inference at the edge of corporate datacenter racks where power and space are constrained
  • Multi-tenant cloud service provider GPU resource pooling leveraging MIG-compatible workload isolation across diverse enterprise tenants
  • Accelerated scientific computing and simulation workloads in life sciences, financial modelling, and engineering design environments

Technical specifications

ManufacturerDell
Manufacturer Part NumberL4-PCIE-24GB
GPU ArchitectureNVIDIA Ada Lovelace
CUDA Cores7,680
Tensor Cores240 (4th Generation)
RT Cores58 (3rd Generation)
GPU Memory24 GB GDDR6
Memory Bandwidth300 GB/s
Memory Interface192-bit
InterfacePCIe 4.0 x16
Form FactorSingle-slot, full-height half-length (FHHL), passive cooling
Thermal Design Power (TDP)72W
Power ConnectorNone (slot-powered)
FP32 Peak Performance30.3 TFLOPS
TF32 Tensor Core Performance120 TFLOPS (sparsity: 242 TFLOPS)
FP16 Tensor Core Performance242 TFLOPS (sparsity: 485 TFLOPS)
INT8 Tensor Core Performance485 TOPS (sparsity: 970 TOPS)
Video Encode Engines2x 4th Generation NVENC
Video Decode Engines1x 3rd Generation NVDEC
Compatible Server PlatformsDell PowerEdge R760, Dell PowerEdge R650
Operating System SupportWindows Server, Red Hat Enterprise Linux, Ubuntu Server, VMware vSphere (with NVIDIA vGPU software)
NVIDIA Software EcosystemCUDA 12.x, TensorRT, Triton Inference Server, NVIDIA AI Enterprise

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandDell
CategoryGPUs
SKUL4-PCIE-24GB
Part NumberL4-PCIE-24GB
ConditionNew
Manufacturer Part NumberL4-PCIE-24GB
GPU ArchitectureNVIDIA Ada Lovelace
CUDA Cores7,680
Tensor Cores240 (4th Generation)
RT Cores58 (3rd Generation)
GPU Memory24 GB GDDR6
Memory Bandwidth300 GB/s
Memory Interface192-bit
InterfacePCIe 4.0 x16
Form FactorSingle-slot, full-height half-length (FHHL), passive cooling
Thermal Design Power (TDP)72W
Power ConnectorNone (slot-powered)
FP32 Peak Performance30.3 TFLOPS
TF32 Tensor Core Performance120 TFLOPS (sparsity: 242 TFLOPS)
FP16 Tensor Core Performance242 TFLOPS (sparsity: 485 TFLOPS)
INT8 Tensor Core Performance485 TOPS (sparsity: 970 TOPS)
Video Encode Engines2x 4th Generation NVENC
Video Decode Engines1x 3rd Generation NVDEC
Compatible Server PlatformsDell PowerEdge R760, Dell PowerEdge R650
Operating System SupportWindows Server, Red Hat Enterprise Linux, Ubuntu Server, VMware vSphere (with NVIDIA vGPU software)
NVIDIA Software EcosystemCUDA 12.x, TensorRT, Triton Inference Server, NVIDIA AI Enterprise

Frequently Asked Questions about NVIDIA L4 24GB GDDR6 PCIe 4.0 – Dell PowerEdge R760 / R650

What server platforms accept the NVIDIA L4 24GB GDDR6 PCIe 4.0 – Dell PowerEdge R760 / R650?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.