NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant

NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant

Brand: HPE | Category: GPUs

SKU: P65892-B21 | Part #: P65892-B21 | MPN: P65892-B21

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant

The NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant (P65892-B21) is built on the NVIDIA Ada Lovelace architecture, delivering a highly versatile accelerator engineered for demanding enterprise AI inference, training, and professional visualization workloads. With 48GB of GDDR6 ECC memory across a 384-bit memory bus, the L40S provides the large memory footprint required for running large language models, generative AI applications, and complex simulation workloads within a standard PCIe form factor — enabling deployment across a wide range of HPE ProLiant server platforms without requiring proprietary interconnects.

The L40S integrates third-generation RT Cores, fourth-generation Tensor Cores, and NVIDIA CUDA cores to support concurrent AI compute and graphics rendering in a single GPU. With a peak FP8 throughput of 1457 TOPS and FP16 Tensor Core performance of 362 TFLOPS, the card is positioned for high-throughput inferencing of transformer-based models alongside 3D rendering and simulation tasks. Its PCIe Gen4 x16 interface ensures broad server compatibility and sufficient bandwidth for enterprise-scale AI and visualization pipelines.

Factory-integrated into HPE's ProLiant ecosystem, this option kit ships with HPE-validated drivers, firmware, and thermal management via an active cooling solution suited to datacenter rack environments. The GPU supports NVIDIA MIG (Multi-Instance GPU) is not applicable to the L40S; instead it supports NVIDIA MPS and time-slicing for workload consolidation. Designed for UAE, GCC, EMEA, and APAC enterprise datacenters, the P65892-B21 is available through Omnixon Global for organizations accelerating AI infrastructure and high-performance compute deployments.

Ideal for

  • Large language model (LLM) inference serving for enterprise AI applications requiring high throughput and large VRAM capacity
  • Generative AI workloads including image synthesis, video generation, and multimodal model deployment in production datacenters
  • Professional 3D visualization and real-time rendering for engineering simulation, digital twin, and media production environments
  • High-performance computing (HPC) workloads such as computational fluid dynamics, molecular dynamics, and scientific modeling on HPE ProLiant platforms
  • AI-assisted medical imaging analysis and healthcare analytics requiring GPU-accelerated inference with large model support
  • Virtual workstation GPU acceleration for remote design, CAD, and visualization via GPU sharing across enterprise user pools

Technical specifications

ManufacturerHPE
Manufacturer Part NumberP65892-B21
GPU ModelNVIDIA L40S
GPU ArchitectureNVIDIA Ada Lovelace
Memory Capacity48GB GDDR6 ECC
Memory Bus Width384-bit
Memory Bandwidth864 GB/s
CUDA Cores18176
FP32 Performance91.6 TFLOPS
FP16 Tensor Core Performance362.05 TFLOPS
FP8 Tensor Core Performance1457.9 TOPS
InterfacePCIe Gen4 x16
Form FactorDual-slot, full-height full-length (FHFL)
CoolingActive (forced air)
Thermal Design Power (TDP)350W
Display OutputsNone (compute and visualization, no display connectors)
NVLink SupportNot supported
ECC MemoryYes
Compatible PlatformsHPE ProLiant servers (validated option kit)
Operating System SupportWindows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server (HPE-validated)

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandHPE
CategoryGPUs
SKUP65892-B21
Part NumberP65892-B21
ConditionNew
Manufacturer Part NumberP65892-B21
GPU ModelNVIDIA L40S
GPU ArchitectureNVIDIA Ada Lovelace
Memory Capacity48GB GDDR6 ECC
Memory Bus Width384-bit
Memory Bandwidth864 GB/s
CUDA Cores18176
FP32 Performance91.6 TFLOPS
FP16 Tensor Core Performance362.05 TFLOPS
FP8 Tensor Core Performance1457.9 TOPS
InterfacePCIe Gen4 x16
Form FactorDual-slot, full-height full-length (FHFL)
CoolingActive (forced air)
Thermal Design Power (TDP)350W
Display OutputsNone (compute and visualization, no display connectors)
NVLink SupportNot supported
ECC MemoryYes
Compatible PlatformsHPE ProLiant servers (validated option kit)
Operating System SupportWindows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server (HPE-validated)

Frequently Asked Questions about NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant

What server platforms accept the NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.