Gigabyte AMD Instinct MI300X 192GB HBM3 OAM GPU Accelerator

Gigabyte AMD Instinct MI300X 192GB HBM3 OAM GPU Accelerator

Brand: Gigabyte | Category: GPUs

SKU: GV-AMI300X-192G | Part #: GV-AMI300X-192G | MPN: GV-AMI300X-192G

Contact for Pricing — Request a Quote

Request a Quote Contact Us

About the Gigabyte AMD Instinct MI300X 192GB HBM3 OAM GPU Accelerator

The Gigabyte AMD Instinct MI300X 192GB HBM3 OAM GPU Accelerator (GV-AMI300X-192G) is built on AMD's CDNA 3 architecture, integrating a unified GPU compute die and HBM3 memory stack within a single OAM (OCP Accelerator Module) form factor. The MI300X APU features 192GB of HBM3 memory across eight stacks with an aggregate memory bandwidth of 5.3 TB/s, making it the highest memory capacity and bandwidth configuration available in its class at launch. This architecture eliminates the CPU-GPU memory boundary by unifying compute and memory on a single package, enabling massive model weights to reside entirely on-accelerator without offloading.

Designed specifically for large-scale AI inference and training, the GV-AMI300X-192G delivers up to 1,307.4 TFLOPS of FP8 compute performance and 383.9 TFLOPS at BF16, supporting the full ROCm software ecosystem including PyTorch, TensorFlow, and JAX. The OAM mechanical standard ensures compatibility with OCP-compliant UBB (Universal Baseboard) server platforms from major hyperscale and enterprise ODM vendors, enabling dense multi-accelerator node configurations with coherent high-speed interconnect via AMD Infinity Fabric.

For enterprise datacenter operators targeting generative AI, large language model (LLM) serving, and high-performance computing workloads, the GV-AMI300X-192G provides a memory footprint sufficient to run 70-billion-parameter and larger transformer models in full precision without tensor parallelism across multiple devices. The accelerator is suited for deployment in AI factories, national research computing clusters, and enterprise private cloud environments where throughput per rack unit, memory capacity, and software ecosystem openness are primary selection criteria.

Ideal for

  • Large language model (LLM) inference serving for models exceeding 70B parameters, fitting full model weights in 192GB on a single accelerator without multi-device tensor parallelism
  • Generative AI training runs including foundation model pre-training and fine-tuning at scale within OCP-compliant multi-accelerator server nodes
  • High-performance scientific computing and simulation workloads in national labs, research institutions, and engineering environments requiring high memory bandwidth and FP64 throughput
  • Enterprise private AI cloud build-outs where open-ecosystem ROCm software compatibility and avoidance of vendor lock-in are architectural requirements
  • Multimodal AI and diffusion model inference requiring large activation memory and sustained high-bandwidth memory access patterns
  • Data center consolidation projects targeting maximum AI compute density per OAM slot within existing OCP Universal Baseboard infrastructure

Technical specifications

ManufacturerGigabyte
Manufacturer Part NumberGV-AMI300X-192G
GPU ModelAMD Instinct MI300X
ArchitectureAMD CDNA 3
Form FactorOAM (OCP Accelerator Module)
Total Memory192 GB HBM3
Memory Stacks8x HBM3
Memory Bandwidth5.3 TB/s
Peak FP8 Compute1,307.4 TFLOPS
Peak FP16 / BF16 Compute383.9 TFLOPS
Peak FP32 Compute163.4 TFLOPS
Peak FP64 Compute81.7 TFLOPS
Compute Units304 CDNA 3 Compute Units
Stream Processors19,456
InterconnectAMD Infinity Fabric (7th Generation)
TDP (Thermal Design Power)750 W
Interface StandardOCP OAM (Universal Baseboard compatible)
Software EcosystemAMD ROCm (PyTorch, TensorFlow, JAX, HIP)
Target PlatformOCP-compliant UBB server nodes

Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.

Technical Specifications

BrandGigabyte
CategoryGPUs
SKUGV-AMI300X-192G
Part NumberGV-AMI300X-192G
ConditionNew
Manufacturer Part NumberGV-AMI300X-192G
GPU ModelAMD Instinct MI300X
ArchitectureAMD CDNA 3
Form FactorOAM (OCP Accelerator Module)
Total Memory192 GB HBM3
Memory Stacks8x HBM3
Memory Bandwidth5.3 TB/s
Peak FP8 Compute1,307.4 TFLOPS
Peak FP16 / BF16 Compute383.9 TFLOPS
Peak FP32 Compute163.4 TFLOPS
Peak FP64 Compute81.7 TFLOPS
Compute Units304 CDNA 3 Compute Units
Stream Processors19,456
InterconnectAMD Infinity Fabric (7th Generation)
TDP (Thermal Design Power)750 W
Interface StandardOCP OAM (Universal Baseboard compatible)
Software EcosystemAMD ROCm (PyTorch, TensorFlow, JAX, HIP)
Target PlatformOCP-compliant UBB server nodes

Frequently Asked Questions about Gigabyte AMD Instinct MI300X 192GB HBM3 OAM GPU Accelerator

What server platforms accept the Gigabyte AMD Instinct MI300X 192GB HBM3 OAM GPU Accelerator?

Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.

How long is the lead time on AI GPUs?

Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.

Do you supply matched networking (Quantum InfiniBand / Spectrum-X)?

Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.

Can you help with NVIDIA AI Enterprise licensing?

Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.