Brand: Supermicro | Category: GPUs
SKU: 100-300000025 | Part #: 100-300000025 | MPN: 100-300000025
Contact for Pricing — Request a Quote
The AMD Instinct MI325X accelerator, offered through Supermicro's validated hardware portfolio under part number 100-300000025, is built on AMD's CDNA 3 compute architecture and represents AMD's highest-memory-capacity GPU accelerator to date. Featuring 256 GB of HBM3E high-bandwidth memory across eight stacks, the MI325X delivers up to 6.0 TB/s of aggregate memory bandwidth, enabling large-scale AI model inference and training workloads that demand both massive memory capacity and extreme data throughput. The accelerator supports FP8, FP16, BF16, FP32, TF32, and INT8 precision formats, with peak matrix compute performance reaching up to 1,307.4 TFLOPS at FP8, making it suited for the most demanding generative AI and scientific computing tasks.
The MI325X is designed for high-density datacenter deployment within Supermicro's OAM (Open Accelerator Module) form factor, enabling integration into multi-accelerator node configurations that scale across large AI infrastructure deployments. Its CDNA 3 architecture incorporates AMD Infinity Fabric technology, supporting coherent multi-GPU communication across nodes and enabling unified memory addressing in multi-chip configurations. The accelerator is optimized for ROCm open-software ecosystem compatibility, providing a flexible path for enterprises deploying PyTorch, TensorFlow, JAX, and HPC frameworks without dependency on proprietary software stacks.
Targeted at enterprise datacenters, cloud service providers, and AI research organizations across the UAE, GCC, EMEA, and APAC regions, the Supermicro-validated MI325X is positioned for deployments requiring the largest available GPU memory footprint in the AMD Instinct line. The 256 GB HBM3E capacity allows entire large language models and recommendation system embeddings to reside on-device, reducing latency and eliminating the need for model sharding across memory boundaries in many practical LLM inference scenarios. Omnixon Global makes this accelerator available to enterprise buyers seeking AMD-based AI infrastructure at scale.
| Manufacturer | Supermicro |
| Manufacturer Part Number | 100-300000025 |
| GPU Model | AMD Instinct MI325X |
| Architecture | AMD CDNA 3 |
| Memory Capacity | 256 GB HBM3E |
| Memory Bandwidth | Up to 6.0 TB/s |
| HBM Stacks | 8 |
| Peak Matrix Performance (FP8) | 1,307.4 TFLOPS |
| Peak Matrix Performance (FP16 / BF16) | 653.7 TFLOPS |
| Peak Compute (FP32) | 163.4 TFLOPS |
| Supported Precision Formats | FP8, FP16, BF16, TF32, FP32, INT8 |
| Form Factor | OAM (Open Accelerator Module) |
| Interconnect | AMD Infinity Fabric |
| Software Ecosystem | AMD ROCm (open-source) |
| Target Deployment | Enterprise datacenter, AI supercomputing, HPC clusters |
| Multi-GPU Scalability | Supported via AMD Infinity Fabric multi-chip coherent interconnect |
| Compatible Frameworks | PyTorch, TensorFlow, JAX, OpenCL, HIP |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Supermicro |
| Category | GPUs |
| SKU | 100-300000025 |
| Part Number | 100-300000025 |
| Condition | New |
| Manufacturer Part Number | 100-300000025 |
| GPU Model | AMD Instinct MI325X |
| Architecture | AMD CDNA 3 |
| Memory Capacity | 256 GB HBM3E |
| Memory Bandwidth | Up to 6.0 TB/s |
| HBM Stacks | 8 |
| Peak Matrix Performance (FP8) | 1,307.4 TFLOPS |
| Peak Matrix Performance (FP16 / BF16) | 653.7 TFLOPS |
| Peak Compute (FP32) | 163.4 TFLOPS |
| Supported Precision Formats | FP8, FP16, BF16, TF32, FP32, INT8 |
| Form Factor | OAM (Open Accelerator Module) |
| Interconnect | AMD Infinity Fabric |
| Software Ecosystem | AMD ROCm (open-source) |
| Target Deployment | Enterprise datacenter, AI supercomputing, HPC clusters |
| Multi-GPU Scalability | Supported via AMD Infinity Fabric multi-chip coherent interconnect |
| Compatible Frameworks | PyTorch, TensorFlow, JAX, OpenCL, HIP |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.