Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU

Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU

Brand: Lenovo | Category: GPUs

SKU: 4X67A72489 | Part #: 4X67A72489 | MPN: 4X67A72489

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU

The Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU (4X67A72489) is a high-performance accelerator built on NVIDIA's Ampere architecture, engineered specifically for demanding enterprise datacenter environments. Delivering 312 teraflops of INT8 inference performance and 77.6 teraflops of FP16 tensor core throughput, the A100 PCIe form factor integrates seamlessly into Lenovo ThinkSystem server platforms without requiring NVLink fabric, making it well-suited for broad datacenter deployment across mixed-workload infrastructure.

At the core of this accelerator is the GA100 GPU die with 6,912 CUDA cores, 432 third-generation Tensor Cores, and 40 GB of High Bandwidth Memory 2 (HBM2) delivering 1,555 GB/s of memory bandwidth. These characteristics make the card exceptionally capable for large-scale deep learning training, scientific simulation, and real-time AI inference across industries including healthcare, financial services, and research computing. The PCIe 4.0 x16 host interface ensures broad platform compatibility while MIG (Multi-Instance GPU) technology allows a single A100 to be partitioned into up to seven isolated GPU instances, enabling efficient multi-tenant workload isolation in virtualized and cloud-native environments.

Validated and factory-integrated by Lenovo for ThinkSystem server platforms, the 4X67A72489 undergoes rigorous compatibility and thermal qualification to meet enterprise reliability standards. The card operates within a 250W TDP thermal envelope with active cooling, and its passive PCIe card design relies on system-level airflow management consistent with Lenovo ThinkSystem rack server chassis specifications. This integration approach simplifies procurement, support, and lifecycle management for enterprise IT teams deploying AI and HPC infrastructure at scale.

Ideal for

  • Large-scale deep learning model training for computer vision, natural language processing, and recommendation systems in enterprise AI platforms
  • High-throughput AI inference serving with multi-instance GPU partitioning to support concurrent, isolated workloads across multiple business units
  • High-performance computing (HPC) workloads including molecular dynamics, computational fluid dynamics, and climate modeling in research and scientific institutions
  • Accelerated data analytics and in-database AI processing for financial services firms requiring low-latency insights from large structured datasets
  • Medical imaging analysis and genomics workloads in healthcare and life sciences environments demanding GPU-accelerated compute with workload isolation
  • Virtualized GPU infrastructure supporting multiple concurrent AI development teams through MIG-enabled multi-tenant GPU partitioning within ThinkSystem server platforms

Technical specifications

ManufacturerLenovo
Manufacturer Part Number4X67A72489
GPU ArchitectureNVIDIA Ampere (GA100)
GPU Memory40 GB HBM2
Memory Bandwidth1,555 GB/s
CUDA Cores6,912
Tensor Cores432 (3rd Generation)
FP64 Performance9.7 TFLOPS
FP32 Performance19.5 TFLOPS
TF32 Tensor Core Performance156 TFLOPS
FP16 Tensor Core Performance312 TFLOPS
INT8 Tensor Core Performance624 TOPS
Host InterfacePCIe 4.0 x16
Form FactorDual-slot, full-height PCIe
Thermal Design Power (TDP)250 W
CoolingPassive (system airflow dependent)
Multi-Instance GPU (MIG)Supported — up to 7 GPU instances
NVLink SupportNot supported (PCIe variant)
ECC MemoryYes
Server CompatibilityLenovo ThinkSystem platforms
Operating System SupportLinux (RHEL, Ubuntu, SLES); Windows Server

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandLenovo
CategoryGPUs
SKU4X67A72489
Part Number4X67A72489
ConditionNew
Manufacturer Part Number4X67A72489
GPU ArchitectureNVIDIA Ampere (GA100)
GPU Memory40 GB HBM2
Memory Bandwidth1,555 GB/s
CUDA Cores6,912
Tensor Cores432 (3rd Generation)
FP64 Performance9.7 TFLOPS
FP32 Performance19.5 TFLOPS
TF32 Tensor Core Performance156 TFLOPS
FP16 Tensor Core Performance312 TFLOPS
INT8 Tensor Core Performance624 TOPS
Host InterfacePCIe 4.0 x16
Form FactorDual-slot, full-height PCIe
Thermal Design Power (TDP)250 W
CoolingPassive (system airflow dependent)
Multi-Instance GPU (MIG)Supported — up to 7 GPU instances
NVLink SupportNot supported (PCIe variant)
ECC MemoryYes
Server CompatibilityLenovo ThinkSystem platforms
Operating System SupportLinux (RHEL, Ubuntu, SLES); Windows Server

Frequently Asked Questions about Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU

What server platforms accept the Lenovo ThinkSystem NVIDIA A100 40GB PCIe GPU?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.