GPU Specifications / AMD

AMD Instinct MI300X

CDNA 3 datacenter GPU with 192 GB of HBM3 memory, 5300 GB/s of bandwidth and up to 2,615 TFLOPS of FP16 tensor compute.

Key Specifications

Memory

192 GB HBM3

Memory Bandwidth

5,300 GB/s

TDP

750 W

Architecture

CDNA 3

Interconnect

Infinity Fabric 3 · 896 GB/s

Est. On-demand Price

~$6.00/h

Hourly rates are indicative on-demand estimates; actual pricing varies by provider and commitment.

Compute Performance

PrecisionPeak throughput
FP64 (double precision)81.7 TFLOPS
FP32 (single precision)163.4 TFLOPS
FP24326.8 TFLOPS
FP16 (tensor)2,615 TFLOPS
INT8 (tensor)5,230 TFLOPS
INT4 (tensor)5,230 TFLOPS

Tensor figures use the vendor's peak numbers (with structured sparsity where supported).

System Requirements

Recommended CPU

AMD EPYC 9554 or Intel Xeon Platinum 8480+

Max VRAM per node (8 GPUs)

1,536 GB

System RAM (min / recommended)

512 / 1024 GB

Minimum PSU

1600 W

LLMs on the Instinct MI300X

Number of Instinct MI300X GPUs needed to serve popular models at 8-bit quantization (including 20% overhead for activations and KV cache).

ModelParamsVRAM (8-bit)GPUs needed
GPT-5.6 Sol2400B2682 GB14x Instinct MI300X
GPT-5 Flagship2100B2347 GB13x Instinct MI300X
GPT-5.6 Luna1400B1565 GB9x Instinct MI300X
Kimi K3 (1.2T)1200B1341 GB7x Instinct MI300X
Kimi K2.6 (1T)1000B1118 GB6x Instinct MI300X
GPT-5.6 Terra800B894 GB5x Instinct MI300X
DeepSeek V4 Pro (671B)671B750 GB4x Instinct MI300X
Llama 4 Behemoth (500B)500B559 GB3x Instinct MI300X
Claude 5 Fable (480B)480B536 GB3x Instinct MI300X
GLM 5.2 (400B)400B447 GB3x Instinct MI300X
Grok 4.5350B391 GB3x Instinct MI300X
Claude 4.8 Opus (300B)300B335 GB2x Instinct MI300X
Muse Spark 1.1300B335 GB2x Instinct MI300X
Grok 4270B302 GB2x Instinct MI300X
Gemini 3.1 Pro250B279 GB2x Instinct MI300X
Qwen 3.7 Max (235B)235B263 GB2x Instinct MI300X
Mistral Large 3 (200B)200B224 GB2x Instinct MI300X
Grok 3 Mini190B212 GB2x Instinct MI300X
Claude 5 Sonnet (175B)175B196 GB2x Instinct MI300X
Gemini 3.5 Flash150B168 GB1x Instinct MI300X
Gemini 2.5 Flash140B156 GB1x Instinct MI300X
Llama 4 Maverick (128B)128B143 GB1x Instinct MI300X
DeepSeek V4 Flash (120B)120B134 GB1x Instinct MI300X
Qwen 3.6 Plus (110B)110B123 GB1x Instinct MI300X
Nova Premier (80B)80B89 GB1x Instinct MI300X
Qwen 3 Coder-Next (80B)80B89 GB1x Instinct MI300X
Claude 4.5 Haiku (70B)70B78 GB1x Instinct MI300X
Llama 3.3 Instruct (70B)70B78 GB1x Instinct MI300X
Mistral Medium 3.5 (70B)70B78 GB1x Instinct MI300X
Yi 1.5 (40B)40B45 GB1x Instinct MI300X
Nova Core (34B)34B38 GB1x Instinct MI300X
DeepSeek V3.1 (32B)32B36 GB1x Instinct MI300X
Gemma 3 (27B)27B30 GB1x Instinct MI300X
Mistral Small 4 (24B)24B27 GB1x Instinct MI300X
Yi 1.5 (15B)15B17 GB1x Instinct MI300X
Phi 4 (14B)14B16 GB1x Instinct MI300X
Nova Lite (12B)12B13 GB1x Instinct MI300X
Llama 3.2 Instruct (11B)11B12 GB1x Instinct MI300X
Gemma 3 (9B)9B10 GB1x Instinct MI300X
Yi 1.5 Lite (9B)9B10 GB1x Instinct MI300X
Phi 4 Mini (7B)7B8 GB1x Instinct MI300X
Phi 3.5 (3.8B)3.8B4 GB1x Instinct MI300X

Frequently Asked Questions

How much VRAM does the AMD Instinct MI300X have?

The AMD Instinct MI300X has 192 GB of HBM3 memory with 5300 GB/s of memory bandwidth.

Which LLMs can run on a single Instinct MI300X?

At 8-bit quantization, a single Instinct MI300X (192 GB) can serve models up to roughly 150B parameters, such as Gemini 3.5 Flash. Larger models require multiple GPUs or more aggressive quantization.

How much does it cost to rent a AMD Instinct MI300X?

On-demand cloud pricing for the Instinct MI300X is around $6.00/hour, i.e. about $4,380/month running 24/7. Actual prices vary by provider, region, and commitment.

What are the power and system requirements of the AMD Instinct MI300X?

The Instinct MI300X has a TDP of 750W. A power supply of at least 1600W per GPU is recommended. Recommended host CPUs: AMD EPYC 9554 or Intel Xeon Platinum 8480+.