Omnixon Global
AMD Instinct MI350X 288GB HBM3E

AMD Instinct MI350X 288GB HBM3E

Brand: Supermicro | Category: GPUs

SKU: 100-300000040 | Part #: 100-300000040 | MPN: 100-300000040

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the AMD Instinct MI350X 288GB HBM3E

The AMD Instinct MI350X 288GB HBM3E is a next-generation compute accelerator built on AMD's CDNA 4 architecture, engineered specifically for the most demanding enterprise AI training, inference, and high-performance computing workloads. Featuring 288GB of HBM3E memory with exceptional memory bandwidth, the MI350X delivers substantial improvements in both compute density and memory capacity over its predecessors, making it a compelling platform for large-scale generative AI model training and deployment at scale.

Supermicro's implementation of the MI350X targets modern datacenter infrastructure where power efficiency, thermal management, and density are critical operational concerns. The accelerator supports AMD's ROCm open software ecosystem, providing compatibility with widely adopted AI and HPC frameworks including PyTorch, TensorFlow, JAX, and OpenMPI-based HPC applications. The MI350X is designed for multi-GPU scale-out environments using AMD Infinity Fabric interconnect technology, enabling coherent high-bandwidth communication across accelerator nodes.

With its large 288GB HBM3E memory footprint, the MI350X is particularly well-suited for inference of frontier large language models (LLMs) and multimodal models that exceed the memory capacity of smaller accelerators, allowing entire model weights to reside on-device without complex sharding strategies. Enterprise datacenter operators across AI cloud services, scientific research institutions, and large-scale model development organizations will find the MI350X a high-capacity, standards-aligned accelerator platform supported through AMD's mature open-source software stack.

Ideal for

  • Large language model (LLM) training and fine-tuning for generative AI applications requiring high memory capacity and compute throughput
  • High-throughput AI inference serving for frontier models including LLMs and multimodal models that benefit from on-device full-weight residence
  • High-performance computing (HPC) simulations in scientific research, computational fluid dynamics, and molecular dynamics workloads
  • Enterprise AI cloud infrastructure buildout requiring dense GPU compute nodes with scale-out interconnect capabilities
  • Data analytics acceleration for large-scale vector database operations and retrieval-augmented generation (RAG) pipelines
  • Government and defense HPC workloads demanding open, auditable software stacks with ROCm-based framework compatibility

Technical specifications

ManufacturerSupermicro
Manufacturer Part Number100-300000040
GPU ArchitectureAMD CDNA 4
GPU ModelAMD Instinct MI350X
Memory Capacity288 GB
Memory TypeHBM3E
Form FactorOAM (Open Accelerator Module)
InterconnectAMD Infinity Fabric
Software EcosystemAMD ROCm (open-source)
Supported FrameworksPyTorch, TensorFlow, JAX, OpenMPI
Target WorkloadsAI Training, AI Inference, HPC
Compute Precision SupportFP64, FP32, FP16, BF16, FP8
Multi-GPU ScalabilityYes, via AMD Infinity Fabric scale-out
CategoryGPU Accelerator

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandSupermicro
CategoryGPUs
SKU100-300000040
Part Number100-300000040
ConditionNew
Manufacturer Part Number100-300000040
GPU ArchitectureAMD CDNA 4
GPU ModelAMD Instinct MI350X
Memory Capacity288 GB
Memory TypeHBM3E
Form FactorOAM (Open Accelerator Module)
InterconnectAMD Infinity Fabric
Software EcosystemAMD ROCm (open-source)
Supported FrameworksPyTorch, TensorFlow, JAX, OpenMPI
Target WorkloadsAI Training, AI Inference, HPC
Compute Precision SupportFP64, FP32, FP16, BF16, FP8
Multi-GPU ScalabilityYes, via AMD Infinity Fabric scale-out

Frequently Asked Questions about AMD Instinct MI350X 288GB HBM3E

What server platforms accept the AMD Instinct MI350X 288GB HBM3E?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.