Brand: HPE | Category: GPUs
SKU: P53867-B21 | Part #: P53867-B21 | MPN: P53867-B21
Contact for Pricing — Request a Quote
The HPE NVIDIA L4 24GB PCIe GPU Accelerator delivers 59.4 TFLOPS at FP32, 118.8 TFLOPS at BF16, and 237.6 TFLOPS at FP8/INT8—the three critical precision points for modern AI inference and machine learning workloads. Built on the NVIDIA Ada Lovelace architecture, this compute-focused accelerator pairs 7,680 CUDA cores with 4th-generation Tensor Cores and 3rd-generation RT Cores to handle demanding inference pipelines at enterprise scale. The 24 GB GDDR6 memory subsystem with 192-bit interface and 300 GB/s memory bandwidth ensures rapid data movement for batch processing scenarios. At just 72 W thermal design power, the HPE NVIDIA L4 (part number P53867-B21) operates efficiently in dense server deployments without excessive cooling overhead.
HPE engineered this full-height, full-length, single-slot accelerator for seamless integration into standard PCIe Gen 4 x16 server slots, eliminating costly infrastructure redesign. Support for Multi-Instance GPU (MIG) partitioning up to 7 instances, NVIDIA vGPU licensing, ECC memory protection, and GPU Direct RDMA capabilities make it ideal for virtualized inference clusters and high-availability production environments. The accelerator operates reliably across a 0°C to 35°C temperature range with no display outputs—a pure compute device engineered for headless deployment. For IT infrastructure teams, network engineers, and AI operations teams evaluating inference acceleration, HPE delivers a power-efficient, standards-based solution that scales across mixed workload clusters. Contact Omnixon Global to request specifications and availability for the HPE NVIDIA L4 24GB PCIe GPU Accelerator (P53867-B21).
| Brand | HPE |
| Category | GPUs |
| SKU | P53867-B21 |
| Part Number | P53867-B21 |
| Condition | New |
| Manufacturer Part Number | P53867-B21 |
| GPU Model | NVIDIA L4 |
| GPU Architecture | NVIDIA Ada Lovelace |
| Memory Capacity | 24 GB GDDR6 |
| Memory Interface | 192-bit |
| Memory Bandwidth | 300 GB/s |
| TDP (Thermal Design Power) | 72 W |
| CUDA Cores | 7680 |
| Tensor Cores | 4th Generation (FP8, FP16, BF16, TF32, INT8) |
| RT Cores | 3rd Generation |
| PCIe Interface | PCIe Gen 4 x16 |
| Form Factor | Full-Height, Full-Length (FHFL), Single-Slot |
| Multi-Instance GPU (MIG) | Supported (up to 7 MIG instances) |
| NVIDIA vGPU Support | Yes |
| NVLink | Not supported |
| Display Outputs | None (compute-only) |
| ECC Memory | Yes |
| GPU Direct RDMA | Supported |
| Operating Temperature | 0°C to 35°C |
| Compatible Server Platform | HPE ProLiant Gen10 Plus / Gen11 (validated) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.