All AI inference chips

MI325X vs MI300X

Spec-sheet and pricing comparison of AMD Instinct MI325X and AMD Instinct MI300X with links to continuously measured LLM inference benchmarks on identical workloads.

Spec-sheet comparison

MI325XMI300XMI325X / MI300X
Memory per chip256 GB HBM3e192 GB HBM31.33x
Memory bandwidth6 TB/s5.3 TB/s1.13x
Dense FP8 compute2,615 TFLOP/s2,615 TFLOP/s1.0x
Dense FP4 computeNot supportedNot supportedn/a
TDP1,000 W750 W1.33x
Hourly rate (neocloud tier)$1.32/hr$1.16/hr1.14x
Scale-up world size8 chips8 chips1.0x

Ratios are spec-sheet values; see the live compare pages for measured deltas.

Frequently asked questions

Which has more memory, MI325X or MI300X?
MI325X offers 256 GB HBM3e per chip versus 192 GB HBM3 on MI300X (1.33x the capacity).
How do MI325X and MI300X prices compare?
At the neocloud tier the SemiAnalysis TCO model rates MI325X at $1.32/hr versus $1.16/hr for MI300X. Hourly price alone is misleading; the per-dollar compare pages divide measured throughput by these rates.
Is MI325X faster than MI300X for LLM inference?
On paper MI325X has 1.0x the dense FP8 compute of MI300X, but delivered tokens per second depend on the model, framework, precision and interactivity target. InferenceX measures both chips daily on identical workloads; see the live compare pages for current results.

See live benchmark results

Every number above is static hardware data. Delivered tokens per second, cost per million tokens and energy per token are measured continuously on the dashboard:

Go deeper with the SemiAnalysis models

InferenceX measures delivered inference performance. The SemiAnalysis institutional models cover the market behind these chips: who ships them, who buys them, and what they cost to own.