Omnixon Global
AMD Instinct MI325X 256GB HBM3E

AMD Instinct MI325X 256GB HBM3E

Brand: Supermicro | Category: GPUs

SKU: 100-300000025 | Part #: 100-300000025 | MPN: 100-300000025

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the AMD Instinct MI325X 256GB HBM3E

The AMD Instinct MI325X accelerator, offered through Supermicro's validated hardware portfolio under part number 100-300000025, is built on AMD's CDNA 3 compute architecture and represents AMD's highest-memory-capacity GPU accelerator to date. Featuring 256 GB of HBM3E high-bandwidth memory across eight stacks, the MI325X delivers up to 6.0 TB/s of aggregate memory bandwidth, enabling large-scale AI model inference and training workloads that demand both massive memory capacity and extreme data throughput. The accelerator supports FP8, FP16, BF16, FP32, TF32, and INT8 precision formats, with peak matrix compute performance reaching up to 1,307.4 TFLOPS at FP8, making it suited for the most demanding generative AI and scientific computing tasks.

The MI325X is designed for high-density datacenter deployment within Supermicro's OAM (Open Accelerator Module) form factor, enabling integration into multi-accelerator node configurations that scale across large AI infrastructure deployments. Its CDNA 3 architecture incorporates AMD Infinity Fabric technology, supporting coherent multi-GPU communication across nodes and enabling unified memory addressing in multi-chip configurations. The accelerator is optimized for ROCm open-software ecosystem compatibility, providing a flexible path for enterprises deploying PyTorch, TensorFlow, JAX, and HPC frameworks without dependency on proprietary software stacks.

Targeted at enterprise datacenters, cloud service providers, and AI research organizations across the UAE, GCC, EMEA, and APAC regions, the Supermicro-validated MI325X is positioned for deployments requiring the largest available GPU memory footprint in the AMD Instinct line. The 256 GB HBM3E capacity allows entire large language models and recommendation system embeddings to reside on-device, reducing latency and eliminating the need for model sharding across memory boundaries in many practical LLM inference scenarios. Omnixon Global makes this accelerator available to enterprise buyers seeking AMD-based AI infrastructure at scale.

Ideal for

  • Large language model (LLM) inference serving, where 256 GB HBM3E capacity allows multi-billion-parameter models to run entirely within a single accelerator's memory space
  • Generative AI training at scale, leveraging FP8 and BF16 matrix compute performance for distributed training jobs across multi-node Supermicro AI server clusters
  • High-performance computing (HPC) simulations in climate modeling, computational fluid dynamics, and molecular dynamics that require both high memory bandwidth and large working-set memory
  • Enterprise recommendation system acceleration, where massive embedding tables benefit directly from the 256 GB on-device memory to avoid costly host-to-device data transfers
  • AI-driven drug discovery and genomics workloads that combine large dataset sizes with complex floating-point compute requirements across sustained multi-day training runs
  • Datacenter AI infrastructure buildouts for sovereign cloud and national AI initiatives across GCC and APAC, where open ROCm software compatibility supports regulatory and vendor-diversity requirements

Technical specifications

ManufacturerSupermicro
Manufacturer Part Number100-300000025
GPU ModelAMD Instinct MI325X
ArchitectureAMD CDNA 3
Memory Capacity256 GB HBM3E
Memory BandwidthUp to 6.0 TB/s
HBM Stacks8
Peak Matrix Performance (FP8)1,307.4 TFLOPS
Peak Matrix Performance (FP16 / BF16)653.7 TFLOPS
Peak Compute (FP32)163.4 TFLOPS
Supported Precision FormatsFP8, FP16, BF16, TF32, FP32, INT8
Form FactorOAM (Open Accelerator Module)
InterconnectAMD Infinity Fabric
Software EcosystemAMD ROCm (open-source)
Target DeploymentEnterprise datacenter, AI supercomputing, HPC clusters
Multi-GPU ScalabilitySupported via AMD Infinity Fabric multi-chip coherent interconnect
Compatible FrameworksPyTorch, TensorFlow, JAX, OpenCL, HIP

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandSupermicro
CategoryGPUs
SKU100-300000025
Part Number100-300000025
ConditionNew
Manufacturer Part Number100-300000025
GPU ModelAMD Instinct MI325X
ArchitectureAMD CDNA 3
Memory Capacity256 GB HBM3E
Memory BandwidthUp to 6.0 TB/s
HBM Stacks8
Peak Matrix Performance (FP8)1,307.4 TFLOPS
Peak Matrix Performance (FP16 / BF16)653.7 TFLOPS
Peak Compute (FP32)163.4 TFLOPS
Supported Precision FormatsFP8, FP16, BF16, TF32, FP32, INT8
Form FactorOAM (Open Accelerator Module)
InterconnectAMD Infinity Fabric
Software EcosystemAMD ROCm (open-source)
Target DeploymentEnterprise datacenter, AI supercomputing, HPC clusters
Multi-GPU ScalabilitySupported via AMD Infinity Fabric multi-chip coherent interconnect
Compatible FrameworksPyTorch, TensorFlow, JAX, OpenCL, HIP

Frequently Asked Questions about AMD Instinct MI325X 256GB HBM3E

What server platforms accept the AMD Instinct MI325X 256GB HBM3E?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.