AMD MI325X

Data center GPU · CDNA3

Summary

The MI325X is an evolution of the MI300X. It has better specs than the NVIDIA H200, but is generally available at lower prices.

Its large VRAM and high memory bandwidth make it especially useful for LLM inference applications. Software support for AMD GPUs has historically been less good than for equivalent NVIDIA GPUs, though this has been improving.

Launched
Q2 2025
VRAM
256 GB
Mem. bandwidth
6,000 GB/s
On-demand from
Compare prices from 1 providers →

Tech specs

VRAM 256 GB HBM3e
Memory bandwidth 6,000 GB/s
Interface PCIe
TDP 1000 W
Supported data types
FP64FP32TF32FP16BF16FP8INT8

Cloud rental prices

1 provider · lowest price per GPU, per hour

Prices updated

On-demand
$2.00 /GPU/h
Provider On-demand Reserved Spot
Vultr $2.00
Compare all AMD MI325X cloud providers & configurations →

Additional providers who may offer AMD MI325X without public pricing: TensorWave, DigitalOcean

Models that fit in VRAM

Open-weights models that fit in VRAM. Estimated using 🤗 accelerate, plus approximation for up to 8K context.

Media

Peak theoretical performance

FP8 Matrix 2,614.9 TFLOPS
INT8 Matrix 2,614.9 TOPS
BF16 Matrix 1,307.4 TFLOPS
FP16 Matrix 1,307.4 TFLOPS
TF32 Matrix 653.7 TFLOPS
FP32 Vector 163.4 TFLOPS
FP32 Matrix 163.4 TFLOPS
FP64 Vector 81.7 TFLOPS
FP64 Matrix 163.4 TFLOPS

Performance figures assume no sparsity; in cases where only sparse performance figures are published by the manufacturer, these are halved to give approximate dense performance.

Resources

Detailed documentation from the manufacturer.

Similar GPUs

Other accelerators you might compare.