Brand: NVIDIA | Category: GPUs
SKU: P52721-B21 | Part #: P52721-B21 | MPN: P52721-B21
Contact for Pricing — Request a Quote
The HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU (P52721-B21) is a purpose-built 2U rack server solution that integrates the NVIDIA L40S data center GPU into HPE's flagship dual-socket ProLiant platform. The NVIDIA L40S is based on the Ada Lovelace architecture and delivers 91.6 TFLOPS of FP32 performance alongside 362 TOPS of INT8 compute, making it one of the most versatile professional GPUs available for enterprise AI inference, 3D rendering, and large-scale simulation workloads. With 48 GB of GDDR6 ECC memory per GPU and a 300W TDP, the L40S is engineered for sustained, high-throughput operation in thermally managed rack environments.
The ProLiant DL380 Gen11 chassis supports up to two NVIDIA L40S GPUs per server node and is powered by 4th Gen Intel Xeon Scalable processors with PCIe Gen 5.0 connectivity, ensuring that CPU-to-GPU data transfer does not become a system bottleneck. HPE iLO 6 out-of-band management, integrated into the DL380 Gen11 platform, enables remote monitoring, firmware orchestration, and health telemetry for the GPU subsystem — a critical capability for large-scale enterprise deployments in distributed data centers across the UAE, GCC, EMEA, and APAC regions.
This integrated solution targets enterprise IT and data center operators requiring a validated, rack-ready system for AI model inference, real-time media processing, virtual workstation delivery via NVIDIA RTX Virtual Workstation (vWS) software, and compute-intensive scientific workloads. The combination of HPE's enterprise server engineering and NVIDIA's Ada Lovelace GPU architecture produces a platform that addresses both current generative AI demands and emerging multi-modal AI inference requirements without sacrificing the operational reliability expected in mission-critical environments.
| Manufacturer | NVIDIA |
| HPE Part Number | P52721-B21 |
| Server Platform | HPE ProLiant DL380 Gen11 |
| Form Factor | 2U Rack Server |
| GPU Model | NVIDIA L40S |
| GPU Architecture | NVIDIA Ada Lovelace |
| GPU Memory | 48 GB GDDR6 ECC |
| GPU Memory Bandwidth | 864 GB/s |
| FP32 Performance | 91.6 TFLOPS |
| TF32 Performance | 183 TFLOPS (sparsity: 366 TFLOPS) |
| FP16 Performance | 362 TFLOPS (sparsity: 724 TFLOPS) |
| INT8 Performance | 724 TOPS (sparsity: 1457 TOPS) |
| GPU TDP | 300W |
| CUDA Cores | 18176 |
| NVENC / NVDEC Engines | 2x NVENC, 2x NVDEC, 1x JPEG |
| PCIe Interface | PCIe Gen 4.0 x16 |
| Server CPU Support | 4th Gen Intel Xeon Scalable Processors |
| Server PCIe Generation | PCIe Gen 5.0 |
| Maximum GPUs per Node | Up to 2 |
| Server Management | HPE iLO 6 |
| Operating System Support | Windows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server, VMware vSphere |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | NVIDIA |
| Category | GPUs |
| SKU | P52721-B21 |
| Part Number | P52721-B21 |
| Condition | New |
| HPE Part Number | P52721-B21 |
| Server Platform | HPE ProLiant DL380 Gen11 |
| Form Factor | 2U Rack Server |
| GPU Model | NVIDIA L40S |
| GPU Architecture | NVIDIA Ada Lovelace |
| GPU Memory | 48 GB GDDR6 ECC |
| GPU Memory Bandwidth | 864 GB/s |
| FP32 Performance | 91.6 TFLOPS |
| TF32 Performance | 183 TFLOPS (sparsity: 366 TFLOPS) |
| FP16 Performance | 362 TFLOPS (sparsity: 724 TFLOPS) |
| INT8 Performance | 724 TOPS (sparsity: 1457 TOPS) |
| GPU TDP | 300W |
| CUDA Cores | 18176 |
| NVENC / NVDEC Engines | 2x NVENC, 2x NVDEC, 1x JPEG |
| PCIe Interface | PCIe Gen 4.0 x16 |
| Server CPU Support | 4th Gen Intel Xeon Scalable Processors |
| Server PCIe Generation | PCIe Gen 5.0 |
| Maximum GPUs per Node | Up to 2 |
| Server Management | HPE iLO 6 |
| Operating System Support | Windows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server, VMware vSphere |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.