Skip to main content
NVIDIA GeForce RTX 4090
NVIDIA · GPU · Ada Lovelace

NVIDIA GeForce RTX 4090 — Benchmarks & Specs

24 GB VRAM450W TDP$1,599 MSRP2022

*Price sourced from Amazon.com. Price and availability subject to change.

Bottom line: how fast is the NVIDIA GeForce RTX 4090?

At 1440p (Ultra), the NVIDIA GeForce RTX 4090 averages 194 fps in Shadow of the Tomb Raider, per HyperCyber. For local LLM inference it generates 2550.0 tokens/sec running llama3.1:8b at FP16 under vllm, per Spheron Blog. In Deep Learning Benchmark (ResNet-50) it scores 128,000 images/sec, per Puget Systems. Its 24 GB of VRAM is the binding constraint for local inference: that capacity fits 32B-parameter models at Q4 without offloading to system RAM.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The NVIDIA GeForce RTX 4090 is a graphics card from the Ada Lovelace family released in 2022 from NVIDIA. Key on-paper specs include 24 GB of GDDR6X VRAM, 450W TDP. It launched with a $1,599 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 4090, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the NVIDIA GeForce RTX 4090 by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Counter-Strike 2 1080p Ultra Native 312 fps 290 fps Phoronix 2023-12-10
Shadow of the Tomb Raider 1440p Ultra 194 fps HyperCyber 2022-10-12
Cyberpunk 2077: Phantom Liberty 1440p Ultra Native 192 fps TechSpot 2023-09-26
Alan Wake 2 1440p Path Traced RT on DLSS Quality 170 fps Tweaktown 2023-10-26
Alan Wake 2 1440p RT Ultra + Path Tracing + DLSS 3.5 FG RT on DLSS 170 fps Tom's Hardware 2023-10-24
Alan Wake 2 1080p Ultra (High, RT Off) Native 154 fps Hardware Times 2024-02-24
007 First Light 4K Ultra DLSS 4.5 145 fps 125 fps PCBench 2025-06-01
Black Myth: Wukong 1440p Cinematic (High quality, RT Off) Native 138 fps TechSpot 2024-08-19
Alan Wake 2 4K Path Traced RT on DLSS Quality 134 fps Tweaktown 2023-10-26
Alan Wake 2 4K RT Ultra + Path Tracing + DLSS 3.5 FG RT on DLSS 134 fps Tom's Hardware 2023-10-24
Alan Wake 2 4K Maximum, Full Ray Tracing RT on DLSS 3.5 134 fps Tom's Hardware 2023-10-24
007 First Light 1080p Ultra 128 fps 115 fps PCBench 2025-06-01
Crimson Desert 1080p Ultra 124 fps 103 fps PCBench 2025-06-01
007 First Light 1440p Ultra 111 fps 91 fps PCBench 2025-06-01

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the NVIDIA GeForce RTX 4090, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
llama3.1:8b FP16 vllm 2550.0 tok/s 18.0 GB Spheron Blog 2026-05-03
qwen3:32b AWQ vllm 650.0 tok/s 22.0 GB Spheron Blog 2026-05-03
llama3:8b ollama 440.0 tok/s LocalLLaMA 2026-04-01
qwen3:30b-moe Q4_K_XL llama.cpp 195.8 tok/s 16.5 GB Hardware Corner 2025-11-06
Llama 2 7B INT4 (AWQ) llama.cpp 194.0 tok/s arXiv (LLM Inference Hardware Survey) 2024-10-06
llama2:7b q4_0 llama.cpp 189.0 tok/s llama.cpp GitHub Discussion #15013 2025-08-01
Llama 3.1 8B Q4_K_M llama.cpp 165.0 tok/s 5.5 GB llama.cpp GitHub 2024-08-22
Llama 3 8B unspecified llama.cpp 150.0 tok/s NVIDIA Developer Blog 2024-10-01
qwen3:32b q4_K_XL llama.cpp 139.7 tok/s Puget Systems 2025-06-01
llama3.1:8b Q4_K_XL llama.cpp 131.0 tok/s 4.8 GB Hardware Corner 2025-11-06
llama3.1:8b q4_K_M llama.cpp 125.0 tok/s 5.5 GB MyAIHardware 2025-05-01
llama3.1:8b q4_K_M llama.cpp 113.0 tok/s 6.2 GB Awesome Agents LLM Leaderboard 2025-06-01

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the NVIDIA GeForce RTX 4090 — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
Deep Learning Benchmark (ResNet-50) 128,000 images/sec Puget Systems 2024-12-10
3DMark Steel Nomad 46,384 points The FPS Review 2024-05-20
3DMark Time Spy Extreme 38,450 pts TechPowerUp 2024-11-02
PassMark G3D Mark 38,066 pts PassMark 2026-04-20
3DMark Time Spy 36,560 points [H]ard|Forum RTX 4090 Time Spy Thread 2023-01-01
3DMark Time Spy GPU 35,200 pts TechPowerUp 2023-11-15
3DMark Port Royal 27,901 points HWBot 2023-01-01
3DMark Port Royal 24,886 points Overclock3D 2022-10-12
3DMark Time Spy Extreme 20,192 points The FPS Review 2022-09-09
3DMark Time Spy Extreme (Graphics) 19,000 pts TechPowerUp 2022-09-01

Products Featuring the NVIDIA GeForce RTX 4090

Full Specifications

tdp w450
vram gb24
vram typeGDDR6X
cuda cores16384
boost clock mhz2520

NVIDIA GeForce RTX 4090 — Frequently Asked Questions

What is the NVIDIA GeForce RTX 4090 best used for?
NVIDIA GeForce RTX 4090 is positioned as a 24 GB VRAM Ada Lovelace-family graphics card. Use it for high-end 4K gaming and local LLM inference. See the synthetic + AI benchmark tables below for measured performance.
When was the NVIDIA GeForce RTX 4090 released, and what was its launch MSRP?
NVIDIA GeForce RTX 4090 launched in 2022 at a $1,599 MSRP. Street prices diverge from launch pricing over a product's lifetime — check the linked Amazon listings on this page for current availability.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the NVIDIA GeForce RTX 4090 run local LLMs?
Yes — NVIDIA GeForce RTX 4090 has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). With 24 GB VRAM, it fits the popular 32B-parameter open-weight models at Q4 quantization comfortably.
Where can I buy the NVIDIA GeForce RTX 4090?
Amazon listings are linked at the top of this page (in the hero CTA) and in the "Products Featuring this Hardware" section below. SpecPicks earns a small affiliate commission on qualifying purchases.

Buying guides that rank the NVIDIA GeForce RTX 4090's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the NVIDIA GeForce RTX 4090

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →
NVIDIA GeForce RTX 4090
NVIDIA GeForce RTX 4090
$2949.99
View on Amazon →