Lenovo ThinkSystem SR675 V3 with NVIDIA L40S 48GB PCIe

Lenovo ThinkSystem SR675 V3 with NVIDIA L40S 48GB PCIe

Brand: Lenovo | Category: GPUs

SKU: 7D9QCTO3WW | Part #: 7D9QCTO3WW | MPN: 7D9QCTO3WW

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Lenovo ThinkSystem SR675 V3 with NVIDIA L40S 48GB PCIe

The Lenovo ThinkSystem SR675 V3 (7D9QCTO3WW) is a high-density 2U dual-socket AMD EPYC-based rack server engineered to support demanding AI inference, professional visualization, and GPU-accelerated compute workloads at enterprise and datacenter scale. Configured with NVIDIA L40S 48GB PCIe GPUs, the SR675 V3 delivers exceptional parallel processing throughput, combining Ada Lovelace GPU architecture with high-capacity GDDR6 memory to handle large model inference, rendering pipelines, and data-intensive analytics simultaneously.

The NVIDIA L40S 48GB PCIe GPU integrated into this configuration is purpose-built for universal GPU workloads, offering a balance of AI training, inference, and professional graphics in a single-slot PCIe form factor. With 142 TFLOPS of FP32 performance, 362 TOPS of INT8 throughput, and 91.6 TFLOPS of TF32 Tensor Core performance, the L40S accelerates generative AI, large language model inference, 3D rendering, and simulation workloads. The 48GB of GDDR6 memory with ECC support enables the loading of large AI models and complex scene data without memory bottlenecks.

The ThinkSystem SR675 V3 chassis is optimized for GPU density, supporting up to eight PCIe Gen 5 GPUs and providing the thermal and power infrastructure necessary to sustain continuous high-performance GPU operation in production datacenter environments. It is compatible with AMD EPYC 9004 series processors, PCIe Gen 5 connectivity, and Lenovo's XClarity management ecosystem, making it a capable platform for organizations across UAE, GCC, EMEA, and APAC regions deploying AI infrastructure, virtual workstation fleets, or HPC clusters.

Ideal for

  • Large language model (LLM) inference serving for enterprise AI applications requiring high throughput and low latency with large model sizes up to and beyond 48GB VRAM per GPU
  • GPU-accelerated 3D rendering and visualization for media, engineering, and design workflows leveraging the L40S Ada Lovelace architecture and RT Cores
  • AI-assisted medical imaging, drug discovery simulation, and genomics workloads requiring high-precision floating-point compute and large memory capacity
  • Virtual workstation delivery at scale using NVIDIA vWS and vPC virtualization, enabling remote professional-grade graphics for distributed enterprise teams
  • Computational fluid dynamics (CFD), finite element analysis (FEA), and scientific simulation workloads that benefit from high FP32 throughput and large GPU memory
  • Generative AI model fine-tuning and multi-model parallel inference in cloud-adjacent on-premises datacenter deployments requiring dense GPU-per-rack configurations

Technical specifications

ManufacturerLenovo
Product LineThinkSystem
ModelSR675 V3
Manufacturer Part Number7D9QCTO3WW
Form Factor2U Rack Server
Processor ArchitectureAMD EPYC 9004 Series (Genoa)
Max Processor Sockets2
GPU ModelNVIDIA L40S 48GB PCIe
GPU ArchitectureAda Lovelace
GPU Memory48GB GDDR6 with ECC
GPU Memory Bandwidth864 GB/s
GPU FP32 Performance91.6 TFLOPS (TF32 Tensor Core), 142 TFLOPS (FP32)
GPU INT8 Tensor Performance362 TOPS
GPU InterfacePCIe Gen 4 x16
Max GPU SupportUp to 8 GPUs (PCIe Gen 5 slots)
PCIe VersionPCIe Gen 5
Max System MemoryUp to 6TB DDR5 ECC RDIMM (24 DIMM slots)
Storage OptionsUp to 24x 2.5" SAS/SATA/NVMe drives
NetworkOnboard 10GbE; optional OCP 3.0 and PCIe NICs
Systems ManagementLenovo XClarity Controller (XCC2), Lenovo XClarity Administrator
Power SupplyRedundant Platinum-rated PSUs
Operating System SupportVMware vSphere, Red Hat Enterprise Linux, Windows Server, Ubuntu

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandLenovo
CategoryGPUs
SKU7D9QCTO3WW
Part Number7D9QCTO3WW
ConditionNew
Product LineThinkSystem
ModelSR675 V3
Manufacturer Part Number7D9QCTO3WW
Form Factor2U Rack Server
Processor ArchitectureAMD EPYC 9004 Series (Genoa)
Max Processor Sockets2
GPU ModelNVIDIA L40S 48GB PCIe
GPU ArchitectureAda Lovelace
GPU Memory48GB GDDR6 with ECC
GPU Memory Bandwidth864 GB/s
GPU FP32 Performance91.6 TFLOPS (TF32 Tensor Core), 142 TFLOPS (FP32)
GPU INT8 Tensor Performance362 TOPS
GPU InterfacePCIe Gen 4 x16
Max GPU SupportUp to 8 GPUs (PCIe Gen 5 slots)
PCIe VersionPCIe Gen 5
Max System MemoryUp to 6TB DDR5 ECC RDIMM (24 DIMM slots)
Storage OptionsUp to 24x 2.5" SAS/SATA/NVMe drives
NetworkOnboard 10GbE; optional OCP 3.0 and PCIe NICs
Systems ManagementLenovo XClarity Controller (XCC2), Lenovo XClarity Administrator
Power SupplyRedundant Platinum-rated PSUs
Operating System SupportVMware vSphere, Red Hat Enterprise Linux, Windows Server, Ubuntu

Frequently Asked Questions about Lenovo ThinkSystem SR675 V3 with NVIDIA L40S 48GB PCIe

What server platforms accept the Lenovo ThinkSystem SR675 V3 with NVIDIA L40S 48GB PCIe?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.