Brand: HPE | Category: GPUs
SKU: P65892-B21 | Part #: P65892-B21 | MPN: P65892-B21
Contact for Pricing — Request a Quote
The NVIDIA L40S 48GB PCIe Gen4 Active Cooling GPU for HPE ProLiant (P65892-B21) is built on the NVIDIA Ada Lovelace architecture, delivering a highly versatile accelerator engineered for demanding enterprise AI inference, training, and professional visualization workloads. With 48GB of GDDR6 ECC memory across a 384-bit memory bus, the L40S provides the large memory footprint required for running large language models, generative AI applications, and complex simulation workloads within a standard PCIe form factor — enabling deployment across a wide range of HPE ProLiant server platforms without requiring proprietary interconnects.
The L40S integrates third-generation RT Cores, fourth-generation Tensor Cores, and NVIDIA CUDA cores to support concurrent AI compute and graphics rendering in a single GPU. With a peak FP8 throughput of 1457 TOPS and FP16 Tensor Core performance of 362 TFLOPS, the card is positioned for high-throughput inferencing of transformer-based models alongside 3D rendering and simulation tasks. Its PCIe Gen4 x16 interface ensures broad server compatibility and sufficient bandwidth for enterprise-scale AI and visualization pipelines.
Factory-integrated into HPE's ProLiant ecosystem, this option kit ships with HPE-validated drivers, firmware, and thermal management via an active cooling solution suited to datacenter rack environments. The GPU supports NVIDIA MIG (Multi-Instance GPU) is not applicable to the L40S; instead it supports NVIDIA MPS and time-slicing for workload consolidation. Designed for UAE, GCC, EMEA, and APAC enterprise datacenters, the P65892-B21 is available through Omnixon Global for organizations accelerating AI infrastructure and high-performance compute deployments.
| Manufacturer | HPE |
| Manufacturer Part Number | P65892-B21 |
| GPU Model | NVIDIA L40S |
| GPU Architecture | NVIDIA Ada Lovelace |
| Memory Capacity | 48GB GDDR6 ECC |
| Memory Bus Width | 384-bit |
| Memory Bandwidth | 864 GB/s |
| CUDA Cores | 18176 |
| FP32 Performance | 91.6 TFLOPS |
| FP16 Tensor Core Performance | 362.05 TFLOPS |
| FP8 Tensor Core Performance | 1457.9 TOPS |
| Interface | PCIe Gen4 x16 |
| Form Factor | Dual-slot, full-height full-length (FHFL) |
| Cooling | Active (forced air) |
| Thermal Design Power (TDP) | 350W |
| Display Outputs | None (compute and visualization, no display connectors) |
| NVLink Support | Not supported |
| ECC Memory | Yes |
| Compatible Platforms | HPE ProLiant servers (validated option kit) |
| Operating System Support | Windows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server (HPE-validated) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P65892-B21 |
| Part Number | P65892-B21 |
| Condition | New |
| Manufacturer Part Number | P65892-B21 |
| GPU Model | NVIDIA L40S |
| GPU Architecture | NVIDIA Ada Lovelace |
| Memory Capacity | 48GB GDDR6 ECC |
| Memory Bus Width | 384-bit |
| Memory Bandwidth | 864 GB/s |
| CUDA Cores | 18176 |
| FP32 Performance | 91.6 TFLOPS |
| FP16 Tensor Core Performance | 362.05 TFLOPS |
| FP8 Tensor Core Performance | 1457.9 TOPS |
| Interface | PCIe Gen4 x16 |
| Form Factor | Dual-slot, full-height full-length (FHFL) |
| Cooling | Active (forced air) |
| Thermal Design Power (TDP) | 350W |
| Display Outputs | None (compute and visualization, no display connectors) |
| NVLink Support | Not supported |
| ECC Memory | Yes |
| Compatible Platforms | HPE ProLiant servers (validated option kit) |
| Operating System Support | Windows Server, Red Hat Enterprise Linux, SUSE Linux Enterprise Server (HPE-validated) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.