Skip to main content
GeForce RTX 4080
NVIDIA · GPU · Ada Lovelace

GeForce RTX 4080 — Benchmarks & Specs

16 GB VRAM320W TDP

*Price sourced from Amazon.com. Price and availability subject to change.

Bottom line: how fast is the GeForce RTX 4080?

At 1440p (Ultra), the GeForce RTX 4080 averages 255 fps in Shadow of the Tomb Raider, per Hardware Times. For local LLM inference it generates 147.5 tokens/sec running qwen3:30b-a3b at IQ3_XXS under llama.cpp, per Glukhov.org — 16GB VRAM LLM benchmarks (llama.cpp). In 3DMark Fire Strike it scores 48,524 points, per Overclocking.com. Its 16 GB of VRAM is the binding constraint for local inference: that capacity fits 13-17B-parameter models at Q4 with room for long context.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The GeForce RTX 4080 is a graphics card from the Ada Lovelace family from NVIDIA. Key on-paper specs include 16 GB of GDDR6X VRAM, 320W TDP. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the GeForce RTX 4080, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the GeForce RTX 4080 by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Shadow of the Tomb Raider 1440p Ultra 255 fps Hardware Times 2024-03-19
Forza Horizon 5 1080p Ultra 241 fps 179 fps DropReference 2026-08-01
F1 2022 1440p Ultra 232 fps Hardware Times 2024-03-19
Hitman 3 1440p Ultra 205 fps Hardware Times 2024-03-19
Horizon Zero Dawn 1440p Ultimate 200 fps KitGuru 2022-11-16
Ghostwire: Tokyo 1440p Ultra 193 fps Hardware Times 2024-03-19
Tiny Tina's Wonderlands 1440p Ultra 183 fps Hardware Times 2024-03-19
Shadow of the Tomb Raider 1440p RT Ultra RT on 180 fps Hardware Times 2024-03-19
Dying Light 2 1440p Ultra 176 fps Hardware Times 2024-03-19
Shadow of the Tomb Raider 4K Ultra RT on DLSS Performance 174 fps Overclocking.com 2022-11-16
Dying Light 2 1440p High 170 fps KitGuru 2022-11-16
Cyberpunk 2077 1440p RT Psycho RT on DLSS Quality 168 fps Hardware Times 2024-03-19
Assassin's Creed Mirage 1080p Ultra 165 fps 122 fps Hardware Times 2023-10-05
F1 2022 4K Ultra 164 fps Hardware Times 2024-03-19

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the GeForce RTX 4080, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
qwen3:30b-a3b IQ3_XXS llama.cpp 147.5 tok/s 13.8 GB Glukhov.org — 16GB VRAM LLM benchmarks (llama.cpp) 2026-03-09
llama2:7b Q4_0 llama.cpp 143.5 tok/s knightli.com 2026-04-23
llama2:7b q4_0 llama.cpp 142.5 tok/s llama.cpp GitHub 2025-08-07
llama2:7b q4_0 llama.cpp 142.5 tok/s llama.cpp GitHub Discussion #15013 2025-08-07
gpt-oss:20b q4_K_M ollama 139.9 tok/s 14.0 GB glukhov.org 2026-03-09
gpt-oss:20b Q4_K_M ollama 139.9 tok/s 14.0 GB Rost Glukhov 2026-03-09
gpt-oss:20b MXFP4 llama.cpp 136.5 tok/s 14.0 GB Hardware Corner 2025-09-01
gemma2:27b IQ4_XS llama.cpp 121.7 tok/s 14.7 GB Glukhov.org — 16GB VRAM LLM benchmarks (llama.cpp) 2026-03-09
llama3.1:8b Q4_K_M ollama 117.0 tok/s 4.9 GB Markaicode 2024-06-01
llama3:8b q4_K_M llama.cpp 106.2 tok/s GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-05-01
qwen3:8b q4_K_M llama.cpp 102.7 tok/s Hardware Corner 2025-09-01
qwen3:8b Q4_K_M llama.cpp 102.7 tok/s Hardware Corner 2025-06-01

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the GeForce RTX 4080 — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
3DMark Fire Strike 48,524 points Overclocking.com 2022-11-15
PassMark G3D Mark 34,433 pts PassMark 2026-04-20
3DMark Time Spy 28,599 points Tom's Hardware 2022-11-03
3DMark Time Spy 26,185 points 3DMark 2023-01-01
3DMark Port Royal 17,663 points 3DMark 2022-12-25
3DMark Port Royal 17,650 points WCCFTech 2022-11-01
3DMark Port Royal 17,600 points HyperCyber 2022-11-03
3DMark Time Spy Extreme 14,178 points WCCFTech 2022-11-01
3DMark Time Spy Extreme 13,977 points Wccftech 2022-10-05
3DMark Speed Way 7,532 points 3DMark 2023-12-25

Products Featuring the GeForce RTX 4080

Full Specifications

rops112
tmus304
tdp w320
opengl4.6
vulkan1.4
directx12 Ultimate (12_2)
foundryTSMC
outputs1x HDMI 2.1 3x DisplayPort 1.4a
vram gb16
gpu chipAD103
l2 cache64 MB
rt cores76
sm count76
bandwidth716.8 GB/s
vram typeGDDR6X
cuda cores9728
generationGeForce 40
process nm5
gpu variantAD103-300-A1
spec sourceTechPowerUp GPU Database

GeForce RTX 4080 — Frequently Asked Questions

What is the GeForce RTX 4080 best used for?
GeForce RTX 4080 is positioned as a 16 GB VRAM Ada Lovelace-family graphics card. Use it for 1440p/4K gaming and mid-tier AI workloads. See the synthetic + AI benchmark tables below for measured performance.
When was the GeForce RTX 4080 released, and what was its launch MSRP?
Release year and launch MSRP aren't on file for GeForce RTX 4080. Current pricing is visible on the linked Amazon product cards below.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the GeForce RTX 4080 run local LLMs?
Yes — GeForce RTX 4080 has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). 16 GB VRAM is enough for 13-17B parameter models at Q4 with comfortable context.
Where can I buy the GeForce RTX 4080?
Amazon listings are linked at the top of this page (in the hero CTA) and in the "Products Featuring this Hardware" section below. SpecPicks earns a small affiliate commission on qualifying purchases.

Buying guides that rank the GeForce RTX 4080's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the GeForce RTX 4080

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →
GeForce RTX 4080
GeForce RTX 4080
$1875.00
View on Amazon →