Skip to main content
NVIDIA GeForce RTX 5090
NVIDIA · GPU · Blackwell

NVIDIA GeForce RTX 5090 — Benchmarks & Specs

32 GB VRAM575W TDP$1,999 MSRP2025

*Price sourced from Amazon.com. Price and availability subject to change.

Bottom line: how fast is the NVIDIA GeForce RTX 5090?

At 1440p (Ultra, Native), the NVIDIA GeForce RTX 5090 averages 350 fps in Final Fantasy XIV with a 281 fps 1% low, per Gamers Nexus. For local LLM inference it generates 5841.0 tokens/sec running Qwen2.5-Coder-7B-Instruct at FP16 under vLLM, per Runpod. In 3DMark Time Spy it scores 46,229 points, per 3DMark. Its 32 GB of VRAM is the binding constraint for local inference: that capacity fits 32B-parameter models at Q4 without offloading to system RAM.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The NVIDIA GeForce RTX 5090 is a graphics card from the Blackwell family released in 2025 from NVIDIA. Key on-paper specs include 32 GB of GDDR7 VRAM, 575W TDP. It launched with a $1,999 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 5090, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the NVIDIA GeForce RTX 5090 by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Final Fantasy XIV 1440p Ultra Native 350 fps 281 fps Gamers Nexus 2025-01-24
Resident Evil 4 1440p Prioritize Graphics 350 fps 281 fps Gamers Nexus 2025-01-30
Final Fantasy XIV: Dawntrail 1440p Maximum 317 fps Gamers Nexus 2025-01-30
Cyberpunk 2077: Phantom Liberty 4K RT Ultra RT on DLSS Quality 286 fps Tom's Hardware 2025-01-24
God of War Ragnarök 1440p Ultra Native 268 fps TechSpot 2025-01-23
Alan Wake 2 4K Ultra RT on 249 fps Tom's Hardware 2025-01-24
Forza Horizon 5 1440p Extreme 235 fps KitGuru 2025-01-30
Resident Evil 4 4K RT Ultra RT on FSR 210 fps Gamers Nexus 2025-01-24
Resident Evil 4 Remake 4K Max RT RT on FSR Quality 210 fps Gamers Nexus 2025-01-24
Resident Evil 4 Remake 4K Prioritize Graphics 207 fps Gamers Nexus 2025-01-24
Resident Evil 4 4K Prioritize Graphics 207 fps Gamers Nexus 2025-01-30
God of War Ragnarök 4K Ultra Native 195 fps TechSpot 2025-01-23
Cyberpunk 2077 1440p Ultra Native 190 fps KitGuru 2025-01-23
Dragon's Dogma 2 1440p Max 189 fps Gamers Nexus 2025-01-24

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the NVIDIA GeForce RTX 5090, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
Qwen2.5-Coder-7B-Instruct FP16 vLLM 5841.0 tok/s Runpod 2025-04-17
Llama 3.1 8B FP8 FP8 vLLM 420.0 tok/s 9.0 GB Runpod 2025-04-30
Llama 2 7B Q4_0 llama.cpp (Vulkan) 263.6 tok/s llama.cpp GitHub 2025-02-15
qwen3-moe:30b q4_K_XL llama.cpp 234.3 tok/s 16.5 GB Hardware Corner 2025-11-06
qwen3moe:30b-a3b q4_K_M llama.cpp 234.3 tok/s 16.5 GB Hardware Corner 2025-06-01
qwen3:8b q4_K_XL llama.cpp 185.9 tok/s 4.8 GB Hardware Corner 2025-11-06
qwen3:8b q4_K_M llama.cpp 185.9 tok/s 4.8 GB Hardware Corner 2025-06-01
llama3.1:8b q4_K_M ollama 149.9 tok/s DatabaseMart 2025-01-30
qwen3:14b q4_K_M llama.cpp 123.8 tok/s 8.5 GB Hardware Corner 2025-06-01
Mixtral 8x7B Q3_K_M llama.cpp 90.0 tok/s 24.0 GB LocalLLaMA 2025-05-02
qwen2.5:14b q4_K_M ollama 89.9 tok/s DatabaseMart 2025-01-30
deepseek-r1:14b q4_K_M ollama 89.1 tok/s DatabaseMart 2025-01-30

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the NVIDIA GeForce RTX 5090 — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
3DMark Time Spy 46,229 points 3DMark 2025-02-01
3DMark Port Royal 45,808 points 3DMark 2025-03-20
PassMark G3D Mark 38,935 pts PassMark 2026-04-20
3DMark Port Royal 36,667 points TechPowerUp 2025-01-22
3DMark Time Spy 32,500 pts TechPowerUp 2025-01-28
3DMark Speed Way 16,532 points 3DMark 2025-03-20
3DMark Speed Way 14,444 points Overclocking.com 2025-01-29
3DMark Steel Nomad (4K) 14,133 points Tom's Hardware 2025-01-22
3DMark Steel Nomad 11,906 points 3DMark 2025-01-24
PassMark G2D Mark 1,412 pts PassMark 2026-04-20

Products Featuring the NVIDIA GeForce RTX 5090

Full Specifications

pciePCIe 5.0 x16
tdp w575
nvlinkfalse
vram gb32
vram typeGDDR7
cuda cores21760
base clock mhz2017
boost clock mhz2407

NVIDIA GeForce RTX 5090 — Frequently Asked Questions

What is the NVIDIA GeForce RTX 5090 best used for?
NVIDIA GeForce RTX 5090 is positioned as a 32 GB VRAM Blackwell-family graphics card. Use it for high-end 4K gaming and local LLM inference. See the synthetic + AI benchmark tables below for measured performance.
When was the NVIDIA GeForce RTX 5090 released, and what was its launch MSRP?
NVIDIA GeForce RTX 5090 launched in 2025 at a $1,999 MSRP. Street prices diverge from launch pricing over a product's lifetime — check the linked Amazon listings on this page for current availability.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the NVIDIA GeForce RTX 5090 run local LLMs?
Yes — NVIDIA GeForce RTX 5090 has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). With 32 GB VRAM, it fits the popular 32B-parameter open-weight models at Q4 quantization comfortably.
Where can I buy the NVIDIA GeForce RTX 5090?
Amazon listings are linked at the top of this page (in the hero CTA) and in the "Products Featuring this Hardware" section below. SpecPicks earns a small affiliate commission on qualifying purchases.

Buying guides that rank the NVIDIA GeForce RTX 5090's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the NVIDIA GeForce RTX 5090

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →
NVIDIA GeForce RTX 5090
NVIDIA GeForce RTX 5090
$5719.99
View on Amazon →