GPU Specifications / AMD
AMD Instinct MI325X
CDNA 3 datacenter GPU with 256 GB of HBM3e memory, 6000 GB/s of bandwidth and up to 2,615 TFLOPS of FP16 tensor compute.
Key Specifications
Memory
256 GB HBM3e
Memory Bandwidth
6,000 GB/s
TDP
1000 W
Architecture
CDNA 3
Interconnect
Infinity Fabric 3 · 896 GB/s
Est. On-demand Price
~$8.00/h
Hourly rates are indicative on-demand estimates; actual pricing varies by provider and commitment.
Compute Performance
| Precision | Peak throughput |
|---|---|
| FP64 (double precision) | 81.7 TFLOPS |
| FP32 (single precision) | 163.4 TFLOPS |
| FP24 | 326.8 TFLOPS |
| FP16 (tensor) | 2,615 TFLOPS |
| INT8 (tensor) | 5,230 TFLOPS |
| INT4 (tensor) | 5,230 TFLOPS |
Tensor figures use the vendor's peak numbers (with structured sparsity where supported).
System Requirements
Recommended CPU
AMD EPYC 9575F or Intel Xeon Platinum 8592+
Max VRAM per node (8 GPUs)
2,048 GB
System RAM (min / recommended)
512 / 1024 GB
Minimum PSU
2000 W
LLMs on the Instinct MI325X
Number of Instinct MI325X GPUs needed to serve popular models at 8-bit quantization (including 20% overhead for activations and KV cache).
| Model | Params | VRAM (8-bit) | GPUs needed |
|---|---|---|---|
| GPT-5.6 Sol | 2400B | 2682 GB | 11x Instinct MI325X |
| GPT-5 Flagship | 2100B | 2347 GB | 10x Instinct MI325X |
| GPT-5.6 Luna | 1400B | 1565 GB | 7x Instinct MI325X |
| Kimi K3 (1.2T) | 1200B | 1341 GB | 6x Instinct MI325X |
| Kimi K2.6 (1T) | 1000B | 1118 GB | 5x Instinct MI325X |
| GPT-5.6 Terra | 800B | 894 GB | 4x Instinct MI325X |
| DeepSeek V4 Pro (671B) | 671B | 750 GB | 3x Instinct MI325X |
| Llama 4 Behemoth (500B) | 500B | 559 GB | 3x Instinct MI325X |
| Claude 5 Fable (480B) | 480B | 536 GB | 3x Instinct MI325X |
| GLM 5.2 (400B) | 400B | 447 GB | 2x Instinct MI325X |
| Grok 4.5 | 350B | 391 GB | 2x Instinct MI325X |
| Claude 4.8 Opus (300B) | 300B | 335 GB | 2x Instinct MI325X |
| Muse Spark 1.1 | 300B | 335 GB | 2x Instinct MI325X |
| Grok 4 | 270B | 302 GB | 2x Instinct MI325X |
| Gemini 3.1 Pro | 250B | 279 GB | 2x Instinct MI325X |
| Qwen 3.7 Max (235B) | 235B | 263 GB | 2x Instinct MI325X |
| Mistral Large 3 (200B) | 200B | 224 GB | 1x Instinct MI325X |
| Grok 3 Mini | 190B | 212 GB | 1x Instinct MI325X |
| Claude 5 Sonnet (175B) | 175B | 196 GB | 1x Instinct MI325X |
| Gemini 3.5 Flash | 150B | 168 GB | 1x Instinct MI325X |
| Gemini 2.5 Flash | 140B | 156 GB | 1x Instinct MI325X |
| Llama 4 Maverick (128B) | 128B | 143 GB | 1x Instinct MI325X |
| DeepSeek V4 Flash (120B) | 120B | 134 GB | 1x Instinct MI325X |
| Qwen 3.6 Plus (110B) | 110B | 123 GB | 1x Instinct MI325X |
| Nova Premier (80B) | 80B | 89 GB | 1x Instinct MI325X |
| Qwen 3 Coder-Next (80B) | 80B | 89 GB | 1x Instinct MI325X |
| Claude 4.5 Haiku (70B) | 70B | 78 GB | 1x Instinct MI325X |
| Llama 3.3 Instruct (70B) | 70B | 78 GB | 1x Instinct MI325X |
| Mistral Medium 3.5 (70B) | 70B | 78 GB | 1x Instinct MI325X |
| Yi 1.5 (40B) | 40B | 45 GB | 1x Instinct MI325X |
| Nova Core (34B) | 34B | 38 GB | 1x Instinct MI325X |
| DeepSeek V3.1 (32B) | 32B | 36 GB | 1x Instinct MI325X |
| Gemma 3 (27B) | 27B | 30 GB | 1x Instinct MI325X |
| Mistral Small 4 (24B) | 24B | 27 GB | 1x Instinct MI325X |
| Yi 1.5 (15B) | 15B | 17 GB | 1x Instinct MI325X |
| Phi 4 (14B) | 14B | 16 GB | 1x Instinct MI325X |
| Nova Lite (12B) | 12B | 13 GB | 1x Instinct MI325X |
| Llama 3.2 Instruct (11B) | 11B | 12 GB | 1x Instinct MI325X |
| Gemma 3 (9B) | 9B | 10 GB | 1x Instinct MI325X |
| Yi 1.5 Lite (9B) | 9B | 10 GB | 1x Instinct MI325X |
| Phi 4 Mini (7B) | 7B | 8 GB | 1x Instinct MI325X |
| Phi 3.5 (3.8B) | 3.8B | 4 GB | 1x Instinct MI325X |
Frequently Asked Questions
How much VRAM does the AMD Instinct MI325X have?
The AMD Instinct MI325X has 256 GB of HBM3e memory with 6000 GB/s of memory bandwidth.
Which LLMs can run on a single Instinct MI325X?
At 8-bit quantization, a single Instinct MI325X (256 GB) can serve models up to roughly 200B parameters, such as Mistral Large 3 (200B). Larger models require multiple GPUs or more aggressive quantization.
How much does it cost to rent a AMD Instinct MI325X?
On-demand cloud pricing for the Instinct MI325X is around $8.00/hour, i.e. about $5,840/month running 24/7. Actual prices vary by provider, region, and commitment.
What are the power and system requirements of the AMD Instinct MI325X?
The Instinct MI325X has a TDP of 1000W. A power supply of at least 2000W per GPU is recommended. Recommended host CPUs: AMD EPYC 9575F or Intel Xeon Platinum 8592+.
Deploy on a GPU cloud
Rent the AMD Instinct MI325X by the hour instead of buying hardware.