Brand: Lenovo | Category: GPUs
SKU: 4X67A72489 | Part #: 4X67A72489 | MPN: 4X67A72489
Contact for Pricing — Request a Quote
The Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU (4X67A72489) is a high-performance accelerator built on NVIDIA's Ampere architecture, engineered specifically for demanding enterprise datacenter environments. Delivering 312 teraflops of INT8 inference performance and 77.6 teraflops of FP16 tensor core throughput, the A100 PCIe form factor integrates seamlessly into Lenovo ThinkSystem server platforms without requiring NVLink fabric, making it well-suited for broad datacenter deployment across mixed-workload infrastructure.
At the core of this accelerator is the GA100 GPU die with 6,912 CUDA cores, 432 third-generation Tensor Cores, and 40 GB of High Bandwidth Memory 2 (HBM2) delivering 1,555 GB/s of memory bandwidth. These characteristics make the card exceptionally capable for large-scale deep learning training, scientific simulation, and real-time AI inference across industries including healthcare, financial services, and research computing. The PCIe 4.0 x16 host interface ensures broad platform compatibility while MIG (Multi-Instance GPU) technology allows a single A100 to be partitioned into up to seven isolated GPU instances, enabling efficient multi-tenant workload isolation in virtualized and cloud-native environments.
Validated and factory-integrated by Lenovo for ThinkSystem server platforms, the 4X67A72489 undergoes rigorous compatibility and thermal qualification to meet enterprise reliability standards. The card operates within a 250W TDP thermal envelope with active cooling, and its passive PCIe card design relies on system-level airflow management consistent with Lenovo ThinkSystem rack server chassis specifications. This integration approach simplifies procurement, support, and lifecycle management for enterprise IT teams deploying AI and HPC infrastructure at scale.
| Manufacturer | Lenovo |
| Manufacturer Part Number | 4X67A72489 |
| GPU Architecture | NVIDIA Ampere (GA100) |
| GPU Memory | 40 GB HBM2 |
| Memory Bandwidth | 1,555 GB/s |
| CUDA Cores | 6,912 |
| Tensor Cores | 432 (3rd Generation) |
| FP64 Performance | 9.7 TFLOPS |
| FP32 Performance | 19.5 TFLOPS |
| TF32 Tensor Core Performance | 156 TFLOPS |
| FP16 Tensor Core Performance | 312 TFLOPS |
| INT8 Tensor Core Performance | 624 TOPS |
| Host Interface | PCIe 4.0 x16 |
| Form Factor | Dual-slot, full-height PCIe |
| Thermal Design Power (TDP) | 250 W |
| Cooling | Passive (system airflow dependent) |
| Multi-Instance GPU (MIG) | Supported — up to 7 GPU instances |
| NVLink Support | Not supported (PCIe variant) |
| ECC Memory | Yes |
| Server Compatibility | Lenovo ThinkSystem platforms |
| Operating System Support | Linux (RHEL, Ubuntu, SLES); Windows Server |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Lenovo |
| Category | GPUs |
| SKU | 4X67A72489 |
| Part Number | 4X67A72489 |
| Condition | New |
| Manufacturer Part Number | 4X67A72489 |
| GPU Architecture | NVIDIA Ampere (GA100) |
| GPU Memory | 40 GB HBM2 |
| Memory Bandwidth | 1,555 GB/s |
| CUDA Cores | 6,912 |
| Tensor Cores | 432 (3rd Generation) |
| FP64 Performance | 9.7 TFLOPS |
| FP32 Performance | 19.5 TFLOPS |
| TF32 Tensor Core Performance | 156 TFLOPS |
| FP16 Tensor Core Performance | 312 TFLOPS |
| INT8 Tensor Core Performance | 624 TOPS |
| Host Interface | PCIe 4.0 x16 |
| Form Factor | Dual-slot, full-height PCIe |
| Thermal Design Power (TDP) | 250 W |
| Cooling | Passive (system airflow dependent) |
| Multi-Instance GPU (MIG) | Supported — up to 7 GPU instances |
| NVLink Support | Not supported (PCIe variant) |
| ECC Memory | Yes |
| Server Compatibility | Lenovo ThinkSystem platforms |
| Operating System Support | Linux (RHEL, Ubuntu, SLES); Windows Server |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.