Skip to main content
AMD Instinct MI300X 192GB
AMD · GPU · CDNA 3

AMD Instinct MI300X 192GB — Benchmarks & Specs

192 GB VRAM750W TDP$15,000 MSRP2023

Bottom line: how fast is the AMD Instinct MI300X 192GB?

For local LLM inference it generates 22021.0 tokens/sec running llama2:70b at FP8 under vllm, per AMD ROCm Blog. In Geekbench 6 OpenCL it scores 379,660 points, per Geekbench Browser. Its 192 GB of VRAM is the binding constraint for local inference: that capacity fits 70B-parameter open-weight models at Q4 without offloading.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The AMD Instinct MI300X 192GB is a graphics card from the CDNA 3 family released in 2023 from AMD. Key on-paper specs include 192 GB of GDDR VRAM, 750W TDP. It launched with a $15,000 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 4 synthetic benchmark results, 12 community AI inference reports, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the AMD Instinct MI300X 192GB, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the AMD Instinct MI300X 192GB, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
llama2:70b FP8 vllm 22021.0 tok/s AMD ROCm Blog 2024-11-01
deepseek-r1:671b FP8 MoAI 21224.6 tok/s Moreh Technical Report 2025-11-13
deepseek-r1:671b vllm 21224.0 tok/s Moreh Technical Report 2025-11-13
llama3.1:70b FP8 vllm 15105.0 tok/s AMD ROCm Blog 2025-01-15
llama3.1:70b FP8 vllm 15105.0 tok/s AMD ROCm Blog 2024-10-30
llama3.1:70b FP8 vllm 10505.0 tok/s AMD ROCm Blog 2025-01-15
deepseek-r1:671b vllm 4574.0 tok/s dstack 2025-02-01
deepseek-r1:671b vllm 4574.0 tok/s dstack Blog 2025-03-01
deepseek-r1:671b vllm 4574.0 tok/s dstack.ai 2025-03-18
llama3.1:405b FP8 vllm 4065.0 tok/s AMD ROCm Blog 2024-10-30
llama3.1:405b FP8 vllm 4065.0 tok/s AMD ROCm Blog 2025-01-15
llama3.1:405b FP8 vllm 3171.0 tok/s AMD ROCm Blog 2025-01-15

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the AMD Instinct MI300X 192GB — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
Geekbench 6 OpenCL 379,660 points Geekbench Browser 2024-06-17
Geekbench 6.3.0 OpenCL 379,660 points Tom's Hardware 2024-06-01
MLPerf Inference v4.1 — Llama-3.1-405B Server (8x MI300X) 24,110 tokens/s AMD ROCm Blog / MLPerf 2025-01-15
MLPerf Inference v4.1 — Llama-2-70B Server (8x MI300X) 22,021 tokens/s AMD ROCm Blog / MLPerf 2025-01-15

Full Specifications

tdp w750
vram gb192
process nm5
bf16 tflops1307
fp16 tflops1307
memory typeHBM3
compute units304
shading units19456
memory bandwidth gbps5300

AMD Instinct MI300X 192GB — Frequently Asked Questions

What is the AMD Instinct MI300X 192GB best used for?
AMD Instinct MI300X 192GB is positioned as a 192 GB VRAM CDNA 3-family graphics card. Use it for high-end 4K gaming and local LLM inference. See the synthetic + AI benchmark tables below for measured performance.
When was the AMD Instinct MI300X 192GB released, and what was its launch MSRP?
AMD Instinct MI300X 192GB launched in 2023 at a $15,000 MSRP. Street prices diverge from launch pricing over a product's lifetime — check the linked Amazon listings on this page for current availability.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the AMD Instinct MI300X 192GB run local LLMs?
Yes — AMD Instinct MI300X 192GB has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). With 192 GB VRAM, it fits the popular 32B-parameter open-weight models at Q4 quantization comfortably.
Where can I buy the AMD Instinct MI300X 192GB?
Active Amazon listings aren't on file for this exact SKU yet. See the linked benchmark sources and the Compare tool for adjacent parts that may be in stock — and check the /benchmarks index for the latest curated picks in this category.

Buying guides that rank the AMD Instinct MI300X 192GB's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the AMD Instinct MI300X 192GB

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →

Hardware benchmark data on SpecPicks

All benchmarks →