Brand: HPE | Category: GPUs
SKU: P38788-B21 | Part #: P38788-B21 | MPN: P38788-B21
Contact for Pricing — Request a Quote
The NVIDIA A30 24GB PCIe Gen4 Passive GPU for HPE ProLiant DL380 (P38788-B21) is a data center-grade accelerator built on NVIDIA's Ampere architecture, engineered specifically for enterprise AI inference, HPC, and data analytics workloads. It delivers 165 TFLOPS of INT8 performance and 10.3 TFLOPS of FP64 tensor performance, supported by 24GB of High Bandwidth Memory 2 (HBM2) with 933 GB/s of memory bandwidth. The passive cooling design integrates seamlessly into the airflow-managed environment of HPE ProLiant DL380 rack servers, making it well-suited for high-density datacenter deployments.
As an HPE factory-integrated option, the A30 is validated and qualified against HPE's ProLiant firmware and system management stack, enabling full visibility through HPE iLO and Integrated Smart Array management. The GPU supports PCIe Gen4 x16 interconnect, providing the bandwidth necessary for latency-sensitive inference pipelines and large-scale model serving. Multi-Instance GPU (MIG) technology allows the A30 to be partitioned into up to four isolated GPU instances, enabling multiple workloads or tenants to share a single physical device with guaranteed quality-of-service.
The A30 supports NVIDIA's full software ecosystem including CUDA, cuDNN, TensorRT, and the NVIDIA AI Enterprise suite, making it a versatile platform for organizations deploying machine learning inference at scale, scientific simulation, and financial modeling. Its passive thermal design and enterprise-grade MTBF characteristics align with the continuous-operation requirements of production datacenter environments across industries including healthcare, financial services, telecommunications, and government.
| Manufacturer | HPE |
| Manufacturer Part Number | P38788-B21 |
| GPU Model | NVIDIA A30 |
| Architecture | NVIDIA Ampere |
| GPU Memory | 24 GB HBM2 |
| Memory Bandwidth | 933 GB/s |
| PCIe Interface | PCIe Gen4 x16 |
| FP32 Performance | 10.3 TFLOPS |
| TF32 Tensor Performance | 82 TFLOPS (sparsity: 165 TFLOPS) |
| FP16 Tensor Performance | 165 TFLOPS (sparsity: 330 TFLOPS) |
| BFLOAT16 Tensor Performance | 165 TFLOPS (sparsity: 330 TFLOPS) |
| INT8 Tensor Performance | 330 TOPS (sparsity: 661 TOPS) |
| CUDA Cores | 3584 |
| Tensor Cores | 224 (3rd Gen) |
| Multi-Instance GPU (MIG) | Up to 4 MIG instances (4x MIG 1g.6gb profile) |
| Thermal Design | Passive (system airflow cooled) |
| TDP (Thermal Design Power) | 165W |
| Form Factor | Full-height, full-length (FHFL) dual-slot |
| Compatible Platform | HPE ProLiant DL380 Gen10 Plus |
| NVLink Support | Not supported |
| ECC Memory | Yes |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | HPE |
| Category | GPUs |
| SKU | P38788-B21 |
| Part Number | P38788-B21 |
| Condition | New |
| Manufacturer Part Number | P38788-B21 |
| GPU Model | NVIDIA A30 |
| Architecture | NVIDIA Ampere |
| GPU Memory | 24 GB HBM2 |
| Memory Bandwidth | 933 GB/s |
| PCIe Interface | PCIe Gen4 x16 |
| FP32 Performance | 10.3 TFLOPS |
| TF32 Tensor Performance | 82 TFLOPS (sparsity: 165 TFLOPS) |
| FP16 Tensor Performance | 165 TFLOPS (sparsity: 330 TFLOPS) |
| BFLOAT16 Tensor Performance | 165 TFLOPS (sparsity: 330 TFLOPS) |
| INT8 Tensor Performance | 330 TOPS (sparsity: 661 TOPS) |
| CUDA Cores | 3584 |
| Tensor Cores | 224 (3rd Gen) |
| Multi-Instance GPU (MIG) | Up to 4 MIG instances (4x MIG 1g.6gb profile) |
| Thermal Design | Passive (system airflow cooled) |
| TDP (Thermal Design Power) | 165W |
| Form Factor | Full-height, full-length (FHFL) dual-slot |
| Compatible Platform | HPE ProLiant DL380 Gen10 Plus |
| NVLink Support | Not supported |
| ECC Memory | Yes |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.