HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU

HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU

Brand: NVIDIA | Category: GPUs

SKU: P52721-B21 | Part #: P52721-B21 | MPN: P52721-B21

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU

The HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU (P52721-B21) is a purpose-built 2U rack server solution that integrates the NVIDIA L40S data center GPU into HPE's flagship dual-socket ProLiant platform. The NVIDIA L40S is based on the Ada Lovelace architecture and delivers 91.6 TFLOPS of FP32 performance alongside 362 TOPS of INT8 compute, making it one of the most versatile professional GPUs available for enterprise AI inference, 3D rendering, and large-scale simulation workloads. With 48 GB of GDDR6 ECC memory per GPU and a 300W TDP, the L40S is engineered for sustained, high-throughput operation in thermally managed rack environments.

The ProLiant DL380 Gen11 chassis supports up to two NVIDIA L40S GPUs per server node and is powered by 4th Gen Intel Xeon Scalable processors with PCIe Gen 5.0 connectivity, ensuring that CPU-to-GPU data transfer does not become a system bottleneck. HPE iLO 6 out-of-band management, integrated into the DL380 Gen11 platform, enables remote monitoring, firmware orchestration, and health telemetry for the GPU subsystem — a critical capability for large-scale enterprise deployments in distributed data centers across the UAE, GCC, EMEA, and APAC regions.

This integrated solution targets enterprise IT and data center operators requiring a validated, rack-ready system for AI model inference, real-time media processing, virtual workstation delivery via NVIDIA RTX Virtual Workstation (vWS) software, and compute-intensive scientific workloads. The combination of HPE's enterprise server engineering and NVIDIA's Ada Lovelace GPU architecture produces a platform that addresses both current generative AI demands and emerging multi-modal AI inference requirements without sacrificing the operational reliability expected in mission-critical environments.

Ideal for

  • Large-scale AI inference serving for generative AI models, large language models (LLMs), and multi-modal AI applications in enterprise data centers
  • High-fidelity 3D visualization, GPU-accelerated rendering, and real-time ray tracing for engineering design, digital twins, and media production pipelines
  • Virtual workstation delivery using NVIDIA RTX Virtual Workstation (vWS) software, enabling GPU-accelerated remote desktops for CAD, BIM, and simulation users
  • Scientific computing and high-performance computing (HPC) simulation workloads requiring sustained FP32 and FP16 throughput in a thermally managed 2U rack footprint
  • Enterprise video transcoding, streaming, and real-time media processing leveraging NVIDIA NVENC and NVDEC hardware encode/decode engines
  • Data center consolidation for organizations migrating from multiple GPU workstation deployments to a centralized, managed server-based GPU compute pool

Technical specifications

ManufacturerNVIDIA
HPE Part NumberP52721-B21
Server PlatformHPE ProLiant DL380 Gen11
Form Factor2U Rack Server
GPU ModelNVIDIA L40S
GPU ArchitectureNVIDIA Ada Lovelace
GPU Memory48 GB GDDR6 ECC
GPU Memory Bandwidth864 GB/s
FP32 Performance91.6 TFLOPS
TF32 Performance183 TFLOPS (sparsity: 366 TFLOPS)
FP16 Performance362 TFLOPS (sparsity: 724 TFLOPS)
INT8 Performance724 TOPS (sparsity: 1457 TOPS)
GPU TDP300W
CUDA Cores18176
NVENC / NVDEC Engines2x NVENC, 2x NVDEC, 1x JPEG
PCIe InterfacePCIe Gen 4.0 x16
Server CPU Support4th Gen Intel Xeon Scalable Processors
Server PCIe GenerationPCIe Gen 5.0
Maximum GPUs per NodeUp to 2
Server ManagementHPE iLO 6
Operating System SupportWindows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server, VMware vSphere

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandNVIDIA
CategoryGPUs
SKUP52721-B21
Part NumberP52721-B21
ConditionNew
HPE Part NumberP52721-B21
Server PlatformHPE ProLiant DL380 Gen11
Form Factor2U Rack Server
GPU ModelNVIDIA L40S
GPU ArchitectureNVIDIA Ada Lovelace
GPU Memory48 GB GDDR6 ECC
GPU Memory Bandwidth864 GB/s
FP32 Performance91.6 TFLOPS
TF32 Performance183 TFLOPS (sparsity: 366 TFLOPS)
FP16 Performance362 TFLOPS (sparsity: 724 TFLOPS)
INT8 Performance724 TOPS (sparsity: 1457 TOPS)
GPU TDP300W
CUDA Cores18176
NVENC / NVDEC Engines2x NVENC, 2x NVDEC, 1x JPEG
PCIe InterfacePCIe Gen 4.0 x16
Server CPU Support4th Gen Intel Xeon Scalable Processors
Server PCIe GenerationPCIe Gen 5.0
Maximum GPUs per NodeUp to 2
Server ManagementHPE iLO 6
Operating System SupportWindows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server, VMware vSphere

Frequently Asked Questions about HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU

What server platforms accept the HPE ProLiant DL380 Gen11 with NVIDIA L40S GPU?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.