Omnixon Global
Supermicro L40S GPU Add-in Board 100-300000070

Supermicro L40S GPU Add-in Board 100-300000070

Brand: Supermicro | Category: GPUs

SKU: 100-300000070 | Part #: 100-300000070 | MPN: 100-300000070

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Supermicro L40S GPU Add-in Board 100-300000070

The Supermicro L40S GPU Add-in Board (part number 100-300000070) is a high-density, full-height full-length PCIe add-in card built around the NVIDIA Ada Lovelace architecture. Featuring 48 GB of GDDR6 ECC memory across a 384-bit memory bus, the L40S delivers exceptional throughput for mixed inference and training workloads, with 91.6 TFLOPS of FP32 performance and support for NVIDIA's third-generation Tensor Cores. The card operates in a single-slot thermal design leveraging active cooling and is optimized for rack-dense enterprise server configurations, including Supermicro's own validated multi-GPU platforms.

The L40S is positioned as a universal data-center GPU, combining professional visualization capabilities with high-performance AI compute on a single card. It supports PCIe Gen 4 x16 connectivity and NVIDIA's full Ada Lovelace feature set including fourth-generation NVENC and NVDEC media engines, enabling simultaneous AI inference and real-time media transcoding without resource contention. The board implements ECC across its full memory capacity, making it suitable for mission-critical enterprise deployments where data integrity under sustained compute load is mandatory.

Designed for large-scale deployment in enterprise data centers across AI inferencing, high-performance computing, and professional visualization, the Supermicro 100-300000070 integrates within standard data-center power and cooling envelopes at a 350 W TDP. Its broad software ecosystem compatibility spans CUDA, TensorRT, NVIDIA AI Enterprise, and major virtualization stacks including vGPU, enabling flexible workload consolidation across bare-metal, virtualized, and containerized environments.

Ideal for

  • Large language model (LLM) and generative AI inference serving at scale, leveraging 48 GB ECC GDDR6 memory to host multi-billion-parameter models on a single card
  • High-throughput video transcoding and AI-accelerated media processing using fourth-generation NVENC/NVDEC engines in broadcast and streaming infrastructure
  • Enterprise GPU virtualization with NVIDIA vGPU software, enabling multiple concurrent virtual workstation or inference instances on a single physical card
  • Scientific and engineering HPC simulation workloads requiring high FP32 and FP64 throughput with memory error correction for computational integrity
  • Professional 3D visualization and CAD/CAM rendering in virtualized or bare-metal workstation environments for design and manufacturing sectors
  • AI-accelerated data analytics and machine learning model training pipelines in on-premises data centers where PCIe Gen 4 bandwidth and large VRAM are critical

Technical specifications

ManufacturerSupermicro
Manufacturer Part Number100-300000070
GPU ArchitectureNVIDIA Ada Lovelace
CUDA Cores18176
Tensor Cores568 (4th Generation)
RT Cores142 (3rd Generation)
Memory Capacity48 GB GDDR6 ECC
Memory Bus Width384-bit
Memory Bandwidth864 GB/s
FP32 Performance91.6 TFLOPS
TF32 Tensor Performance183 TFLOPS (sparsity: 366 TFLOPS)
FP16 Tensor Performance362.05 TFLOPS (sparsity: 724.1 TFLOPS)
INT8 Tensor Performance724.1 TOPS (sparsity: 1457.9 TOPS)
PCIe InterfacePCIe 4.0 x16
Form FactorFull-Height, Full-Length (FHFL)
Thermal Design Power (TDP)350 W
CoolingActive (dual-slot blower fan)
Display Outputs4x DisplayPort 1.4
NVLink SupportNone
GPU VirtualizationNVIDIA vGPU (GRID) supported
ECC Memory SupportYes
Video Encode/Decode Engines2x NVENC (4th Gen), 2x NVDEC (5th Gen)
Operating System SupportLinux, Windows Server

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandSupermicro
CategoryGPUs
SKU100-300000070
Part Number100-300000070
ConditionNew
Manufacturer Part Number100-300000070
GPU ArchitectureNVIDIA Ada Lovelace
CUDA Cores18176
Tensor Cores568 (4th Generation)
RT Cores142 (3rd Generation)
Memory Capacity48 GB GDDR6 ECC
Memory Bus Width384-bit
Memory Bandwidth864 GB/s
FP32 Performance91.6 TFLOPS
TF32 Tensor Performance183 TFLOPS (sparsity: 366 TFLOPS)
FP16 Tensor Performance362.05 TFLOPS (sparsity: 724.1 TFLOPS)
INT8 Tensor Performance724.1 TOPS (sparsity: 1457.9 TOPS)
PCIe InterfacePCIe 4.0 x16
Form FactorFull-Height, Full-Length (FHFL)
Thermal Design Power (TDP)350 W
CoolingActive (dual-slot blower fan)
Display Outputs4x DisplayPort 1.4
NVLink SupportNone
GPU VirtualizationNVIDIA vGPU (GRID) supported
ECC Memory SupportYes
Video Encode/Decode Engines2x NVENC (4th Gen), 2x NVDEC (5th Gen)
Operating System SupportLinux, Windows Server

Frequently Asked Questions about Supermicro L40S GPU Add-in Board 100-300000070

What server platforms accept the Supermicro L40S GPU Add-in Board 100-300000070?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.