Dell PowerEdge XE9680 8x MI300X OAM Server – CTO GPU Node

Dell PowerEdge XE9680 8x MI300X OAM Server – CTO GPU Node

Brand: Dell | Category: GPUs

SKU: XE9680-MI300X-8GPU | Part #: XE9680-MI300X-8GPU | MPN: XE9680-MI300X-8GPU

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Dell PowerEdge XE9680 8x MI300X OAM Server – CTO GPU Node

The Dell PowerEdge XE9680 configured with eight AMD Instinct MI300X OAM accelerators represents one of the most memory-dense and compute-capable GPU server platforms available for enterprise AI and HPC deployments. The XE9680 chassis is purpose-built to house OAM (Open Accelerator Module) form-factor GPUs, and the MI300X variant brings AMD's CDNA 3 architecture to bear with a unified CPU+GPU memory architecture—each MI300X integrates 192 GB of HBM3 memory, yielding a total of 1.5 TB of high-bandwidth accelerator memory across all eight GPUs in a single 5U node. This extraordinarily large unified memory pool enables workloads that would otherwise require model sharding across multiple nodes to run within a single server, dramatically reducing inter-node communication overhead.

The XE9680 platform supports PCIe Gen 5 interconnects and is engineered around high-throughput AMD Infinity Fabric links between accelerators, enabling rapid all-to-all GPU communication suited to large-scale model parallelism. The system accommodates high-core-count AMD EPYC (Genoa-class) processors, high-capacity DDR5 system memory, and NVMe storage expansion, making it a fully self-contained AI training and inference node. Liquid cooling readiness and a dense 5U form factor address the thermal and spatial challenges inherent in deploying this class of accelerator at data center scale.

As a Configure-To-Order (CTO) offering, the XE9680 MI300X GPU Node allows enterprise procurement teams to tailor CPU, DRAM, storage, networking, and firmware configurations to specific workload and infrastructure requirements. This platform is well suited to large language model training, generative AI inference, high-performance scientific computing, and GPU-accelerated data analytics pipelines in UAE, GCC, EMEA, and APAC enterprise and hyperscale data center environments.

Ideal for

  • Large language model (LLM) pre-training and fine-tuning requiring multi-hundred-billion parameter models to reside entirely within on-accelerator HBM3 memory without multi-node sharding
  • Generative AI inference serving for production deployments of foundation models where low-latency, high-throughput responses are required at enterprise scale
  • High-performance scientific computing and simulation workloads in energy, life sciences, and engineering that require both dense compute throughput and large fast-access memory capacity
  • GPU-accelerated data analytics and machine learning pipelines processing large in-memory datasets for financial services, telecommunications, and government sectors
  • Multi-tenant AI-as-a-Service infrastructure within private or sovereign cloud data centers serving GCC and APAC regulated industries
  • Retrieval-augmented generation (RAG) and vector database acceleration where large embedding models and indices must be held in high-bandwidth accelerator memory for real-time query performance

Technical specifications

ManufacturerDell
Manufacturer Part NumberXE9680-MI300X-8GPU
Form Factor5U Rack Server
GPU Accelerators8x AMD Instinct MI300X OAM
GPU ArchitectureAMD CDNA 3
GPU Memory per Accelerator192 GB HBM3
Total Accelerator Memory1.5 TB HBM3 (8 x 192 GB)
GPU InterconnectAMD Infinity Fabric (OAM baseboard)
CPU SupportDual AMD EPYC (Genoa) processors
System Memory TypeDDR5
PCIe GenerationPCIe Gen 5
GPU Module Form FactorOAM (Open Accelerator Module)
Storage InterfaceNVMe (U.2/E3.S)
Network InterfaceConfigurable (CTO); supports high-speed Ethernet and InfiniBand options
Power SupplyHigh-efficiency redundant PSUs (configuration-dependent)
CoolingAir cooling standard; liquid cooling ready
Operating System SupportLinux (RHEL, Ubuntu); ROCm software stack
Configuration ModelCTO (Configure-To-Order)
Target WorkloadsGenerative AI training and inference, LLM, HPC, scientific computing
Regional AvailabilityUAE, GCC, EMEA, APAC

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandDell
CategoryGPUs
SKUXE9680-MI300X-8GPU
Part NumberXE9680-MI300X-8GPU
ConditionNew
Manufacturer Part NumberXE9680-MI300X-8GPU
Form Factor5U Rack Server
GPU Accelerators8x AMD Instinct MI300X OAM
GPU ArchitectureAMD CDNA 3
GPU Memory per Accelerator192 GB HBM3
Total Accelerator Memory1.5 TB HBM3 (8 x 192 GB)
GPU InterconnectAMD Infinity Fabric (OAM baseboard)
CPU SupportDual AMD EPYC (Genoa) processors
System Memory TypeDDR5
PCIe GenerationPCIe Gen 5
GPU Module Form FactorOAM (Open Accelerator Module)
Storage InterfaceNVMe (U.2/E3.S)
Network InterfaceConfigurable (CTO); supports high-speed Ethernet and InfiniBand options
Power SupplyHigh-efficiency redundant PSUs (configuration-dependent)
CoolingAir cooling standard; liquid cooling ready
Operating System SupportLinux (RHEL, Ubuntu); ROCm software stack
Configuration ModelCTO (Configure-To-Order)
Target WorkloadsGenerative AI training and inference, LLM, HPC, scientific computing
Regional AvailabilityUAE, GCC, EMEA, APAC

Frequently Asked Questions about Dell PowerEdge XE9680 8x MI300X OAM Server – CTO GPU Node

What server platforms accept the Dell PowerEdge XE9680 8x MI300X OAM Server – CTO GPU Node?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.