Brand: HPE | Category: GPUs
SKU: P40228-B21 | Part #: P40228-B21 | MPN: P40228-B21
Contact for Pricing — Request a Quote
The NVIDIA A16 64GB PCIe Gen4 GPU (HPE Part P40228-B21) is a quad-chip, high-density graphics and compute accelerator purpose-built for virtualized enterprise workloads running on HPE ProLiant DL380 Gen10 Plus servers. Based on NVIDIA's Ampere architecture, the A16 integrates four GA107 GPU dies on a single PCIe Gen4 card, delivering a combined 64GB of GDDR6 memory (16GB per GPU) and 1,024 CUDA cores per chip. The card is optimized for NVIDIA Virtual PC (vPC), Virtual Applications (vApps), and Virtual Workstation (vWS) profiles under the NVIDIA Virtual GPU (vGPU) software stack, enabling multiple concurrent virtual desktop and application sessions with hardware-level GPU isolation.
Designed for data center density and efficiency, the A16 operates within a 250W TDP envelope and uses a passive cooling design that relies on the server chassis airflow of HPE ProLiant systems. PCIe Gen4 x16 host connectivity provides the bandwidth headroom needed to support concurrent multi-user GPU virtualization without bottlenecking I/O. Each of the four A16 GPU dies supports independent vGPU partitioning, allowing administrators to allocate discrete GPU slices to individual virtual machines with guaranteed frame buffer and compute resources.
As an HPE factory-integrated option (HPE option kit P40228-B21), this GPU is validated and supported within the HPE ProLiant DL380 ecosystem, ensuring compatibility with HPE's server management tools, Intelligent Provisioning, and Integrated Lights-Out (iLO). It is particularly well-suited for enterprise IT environments in industries such as healthcare, financial services, oil and gas, and engineering that require scalable, secure, multi-user GPU-accelerated virtual desktops and remote workstation experiences across distributed or centralized data centers.
| Manufacturer | HPE |
| Manufacturer Part Number | P40228-B21 |
| GPU Model | NVIDIA A16 |
| GPU Architecture | NVIDIA Ampere (GA107 x4) |
| Number of GPU Dies per Card | 4 |
| Total GPU Memory | 64GB GDDR6 |
| Memory per GPU Die | 16GB GDDR6 |
| CUDA Cores per GPU Die | 1,024 |
| Total CUDA Cores | 4,096 |
| Host Interface | PCIe Gen4 x16 |
| Thermal Design Power (TDP) | 250W |
| Cooling | Passive (chassis airflow dependent) |
| Form Factor | Full-height, full-length (FHFL) single-slot |
| Display Outputs | None (compute and virtualization only) |
| vGPU Software Support | NVIDIA Virtual PC (vPC), Virtual Applications (vApps), Virtual Workstation (vWS) |
| Compatible Server Platform | HPE ProLiant DL380 Gen10 Plus |
| ECC Memory Support | Yes |
| API Support | DirectX 12, OpenGL 4.6, Vulkan, CUDA, OpenCL |
| NVLink / NVSwitch | Not supported |
| Product Type | HPE Option Kit |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P40228-B21 |
| Part Number | P40228-B21 |
| Condition | New |
| Manufacturer Part Number | P40228-B21 |
| GPU Model | NVIDIA A16 |
| GPU Architecture | NVIDIA Ampere (GA107 x4) |
| Number of GPU Dies per Card | 4 |
| Total GPU Memory | 64GB GDDR6 |
| Memory per GPU Die | 16GB GDDR6 |
| CUDA Cores per GPU Die | 1,024 |
| Total CUDA Cores | 4,096 |
| Host Interface | PCIe Gen4 x16 |
| Thermal Design Power (TDP) | 250W |
| Cooling | Passive (chassis airflow dependent) |
| Form Factor | Full-height, full-length (FHFL) single-slot |
| Display Outputs | None (compute and virtualization only) |
| vGPU Software Support | NVIDIA Virtual PC (vPC), Virtual Applications (vApps), Virtual Workstation (vWS) |
| Compatible Server Platform | HPE ProLiant DL380 Gen10 Plus |
| ECC Memory Support | Yes |
| API Support | DirectX 12, OpenGL 4.6, Vulkan, CUDA, OpenCL |
| NVLink / NVSwitch | Not supported |
| Product Type | HPE Option Kit |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.