Brand: Lenovo | Category: GPUs
SKU: 7D76CTO2WW | Part #: 7D76CTO2WW | MPN: 7D76CTO2WW
Contact for Pricing — Request a Quote
The Lenovo ThinkSystem SR650 V3 (7D76CTO2WW) is a 2U dual-socket rack server engineered for demanding enterprise, datacenter, and AI inference workloads. Built on the Intel Xeon Scalable 4th Generation (Sapphire Rapids) processor platform, the SR650 V3 supports up to two processors and delivers a high-bandwidth memory and I/O architecture optimized for GPU-accelerated computing. When configured with the NVIDIA L4 24GB PCIe GPU, the system targets a broad spectrum of AI and accelerated computing tasks, combining the stability and manageability of Lenovo's ThinkSystem server portfolio with NVIDIA's Ada Lovelace generation inference silicon.
The NVIDIA L4 24GB PCIe adapter installed in this configuration is a 72-watt single-slot GPU based on the NVIDIA Ada Lovelace architecture, featuring 24 GB of GDDR6 ECC memory and 58.9 TOPS of INT8 performance. The L4 is purpose-built for energy-efficient AI inference, video transcoding, and professional visualization in space- and power-constrained datacenter environments. Its low TDP makes it compatible with standard PCIe slots without requiring supplemental power connectors, enabling high-density GPU deployments within the SR650 V3's multiple PCIe Gen 5 expansion slots. The combination supports NVIDIA's full software ecosystem including CUDA, TensorRT, and NVIDIA AI Enterprise.
The ThinkSystem SR650 V3 platform supports up to 32 DIMM slots accommodating DDR5 memory at speeds up to 4800 MT/s, PCIe Gen 5 connectivity, and Lenovo's XClarity management suite for lifecycle, firmware, and telemetry management. The system is designed to integrate into existing enterprise IT environments with support for Lenovo Neptune liquid cooling options and is well suited for organizations across UAE, GCC, EMEA, and APAC regions seeking scalable, GPU-accelerated rack infrastructure for production AI and data-intensive operations.
| Manufacturer | Lenovo |
| Manufacturer Part Number | 7D76CTO2WW |
| Product Line | ThinkSystem SR650 V3 |
| Form Factor | 2U Rack Server |
| Processor Support | Dual Intel Xeon Scalable 4th Generation (Sapphire Rapids), up to 60 cores per socket |
| Memory Slots | 32 x DIMM slots |
| Memory Type | DDR5 |
| Memory Speed | Up to 4800 MT/s |
| PCIe Generation | PCIe Gen 5 |
| GPU Model | NVIDIA L4 24GB PCIe |
| GPU Architecture | NVIDIA Ada Lovelace |
| GPU Memory | 24 GB GDDR6 ECC |
| GPU TDP | 72 W |
| GPU INT8 Performance | 58.9 TOPS |
| GPU Form Factor | Single-slot, full-height PCIe (no auxiliary power connector required) |
| GPU CUDA Cores | 7680 |
| GPU Tensor Cores | 240 (4th Generation) |
| Storage Bays | Up to 24 x 2.5-inch SAS/SATA/NVMe drive bays (configuration dependent) |
| Network | Onboard Broadcom 1GbE dual-port, additional OCP 3.0 and PCIe NIC options |
| Management | Lenovo XClarity Controller 2 (XCC2), XClarity Administrator support |
| Cooling Options | Air cooling standard; Lenovo Neptune direct water cooling optional |
| Power Supply | Redundant Platinum or Titanium hot-swap PSU options |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Lenovo |
| Category | GPUs |
| SKU | 7D76CTO2WW |
| Part Number | 7D76CTO2WW |
| Condition | New |
| Manufacturer Part Number | 7D76CTO2WW |
| Product Line | ThinkSystem SR650 V3 |
| Form Factor | 2U Rack Server |
| Processor Support | Dual Intel Xeon Scalable 4th Generation (Sapphire Rapids), up to 60 cores per socket |
| Memory Slots | 32 x DIMM slots |
| Memory Type | DDR5 |
| Memory Speed | Up to 4800 MT/s |
| PCIe Generation | PCIe Gen 5 |
| GPU Model | NVIDIA L4 24GB PCIe |
| GPU Architecture | NVIDIA Ada Lovelace |
| GPU Memory | 24 GB GDDR6 ECC |
| GPU TDP | 72 W |
| GPU INT8 Performance | 58.9 TOPS |
| GPU Form Factor | Single-slot, full-height PCIe (no auxiliary power connector required) |
| GPU CUDA Cores | 7680 |
| GPU Tensor Cores | 240 (4th Generation) |
| Storage Bays | Up to 24 x 2.5-inch SAS/SATA/NVMe drive bays (configuration dependent) |
| Network | Onboard Broadcom 1GbE dual-port, additional OCP 3.0 and PCIe NIC options |
| Management | Lenovo XClarity Controller 2 (XCC2), XClarity Administrator support |
| Cooling Options | Air cooling standard; Lenovo Neptune direct water cooling optional |
| Power Supply | Redundant Platinum or Titanium hot-swap PSU options |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.