NVIDIA GeForce RTX 5090 — Benchmarks & Specs
*Price sourced from Amazon.com. Price and availability subject to change.
Bottom line: how fast is the NVIDIA GeForce RTX 5090?
At 1440p (Ultra, Native), the NVIDIA GeForce RTX 5090 averages 350 fps in Final Fantasy XIV with a 281 fps 1% low, per Gamers Nexus. For local LLM inference it generates 5841.0 tokens/sec running Qwen2.5-Coder-7B-Instruct at FP16 under vLLM, per Runpod. In 3DMark Time Spy it scores 46,229 points, per 3DMark. Its 32 GB of VRAM is the binding constraint for local inference: that capacity fits 32B-parameter models at Q4 without offloading to system RAM.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The NVIDIA GeForce RTX 5090 is a graphics card from the Blackwell family released in 2025 from NVIDIA. Key on-paper specs include 32 GB of GDDR7 VRAM, 575W TDP. It launched with a $1,999 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 5090, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).
Gaming Performance (measured FPS)
Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.
| Game | Resolution | Settings | Relative | Avg FPS | 1% low | Source |
|---|---|---|---|---|---|---|
| Final Fantasy XIV | 1440p | Ultra Native | 350 fps | 281 fps | Gamers Nexus 2025-01-24 | |
| Resident Evil 4 | 1440p | Prioritize Graphics | 350 fps | 281 fps | Gamers Nexus 2025-01-30 | |
| Final Fantasy XIV: Dawntrail | 1440p | Maximum | 317 fps | — | Gamers Nexus 2025-01-30 | |
| Cyberpunk 2077: Phantom Liberty | 4K | RT Ultra RT on DLSS Quality | 286 fps | — | Tom's Hardware 2025-01-24 | |
| God of War Ragnarök | 1440p | Ultra Native | 268 fps | — | TechSpot 2025-01-23 | |
| Alan Wake 2 | 4K | Ultra RT on | 249 fps | — | Tom's Hardware 2025-01-24 | |
| Forza Horizon 5 | 1440p | Extreme | 235 fps | — | KitGuru 2025-01-30 | |
| Resident Evil 4 | 4K | RT Ultra RT on FSR | 210 fps | — | Gamers Nexus 2025-01-24 | |
| Resident Evil 4 Remake | 4K | Max RT RT on FSR Quality | 210 fps | — | Gamers Nexus 2025-01-24 | |
| Resident Evil 4 Remake | 4K | Prioritize Graphics | 207 fps | — | Gamers Nexus 2025-01-24 | |
| Resident Evil 4 | 4K | Prioritize Graphics | 207 fps | — | Gamers Nexus 2025-01-30 | |
| God of War Ragnarök | 4K | Ultra Native | 195 fps | — | TechSpot 2025-01-23 | |
| Cyberpunk 2077 | 1440p | Ultra Native | 190 fps | — | KitGuru 2025-01-23 | |
| Dragon's Dogma 2 | 1440p | Max | 189 fps | — | Gamers Nexus 2025-01-24 |
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| Qwen2.5-Coder-7B-Instruct | FP16 vLLM | 5841.0 tok/s | — | Runpod 2025-04-17 | |
| Llama 3.1 8B FP8 | FP8 vLLM | 420.0 tok/s | 9.0 GB | Runpod 2025-04-30 | |
| Llama 2 7B | Q4_0 llama.cpp (Vulkan) | 263.6 tok/s | — | llama.cpp GitHub 2025-02-15 | |
| qwen3-moe:30b | q4_K_XL llama.cpp | 234.3 tok/s | 16.5 GB | Hardware Corner 2025-11-06 | |
| qwen3moe:30b-a3b | q4_K_M llama.cpp | 234.3 tok/s | 16.5 GB | Hardware Corner 2025-06-01 | |
| qwen3:8b | q4_K_XL llama.cpp | 185.9 tok/s | 4.8 GB | Hardware Corner 2025-11-06 | |
| qwen3:8b | q4_K_M llama.cpp | 185.9 tok/s | 4.8 GB | Hardware Corner 2025-06-01 | |
| llama3.1:8b | q4_K_M ollama | 149.9 tok/s | — | DatabaseMart 2025-01-30 | |
| qwen3:14b | q4_K_M llama.cpp | 123.8 tok/s | 8.5 GB | Hardware Corner 2025-06-01 | |
| Mixtral 8x7B | Q3_K_M llama.cpp | 90.0 tok/s | 24.0 GB | LocalLLaMA 2025-05-02 | |
| qwen2.5:14b | q4_K_M ollama | 89.9 tok/s | — | DatabaseMart 2025-01-30 | |
| deepseek-r1:14b | q4_K_M ollama | 89.1 tok/s | — | DatabaseMart 2025-01-30 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| 3DMark Time Spy | 46,229 points | 3DMark 2025-02-01 | |
| 3DMark Port Royal | 45,808 points | 3DMark 2025-03-20 | |
| PassMark G3D Mark | 38,935 pts | PassMark 2026-04-20 | |
| 3DMark Port Royal | 36,667 points | TechPowerUp 2025-01-22 | |
| 3DMark Time Spy | 32,500 pts | TechPowerUp 2025-01-28 | |
| 3DMark Speed Way | 16,532 points | 3DMark 2025-03-20 | |
| 3DMark Speed Way | 14,444 points | Overclocking.com 2025-01-29 | |
| 3DMark Steel Nomad (4K) | 14,133 points | Tom's Hardware 2025-01-22 | |
| 3DMark Steel Nomad | 11,906 points | 3DMark 2025-01-24 | |
| PassMark G2D Mark | 1,412 pts | PassMark 2026-04-20 |
Products Featuring the NVIDIA GeForce RTX 5090
Full Specifications
| pcie | PCIe 5.0 x16 |
|---|---|
| tdp w | 575 |
| nvlink | false |
| vram gb | 32 |
| vram type | GDDR7 |
| cuda cores | 21760 |
| base clock mhz | 2017 |
| boost clock mhz | 2407 |
NVIDIA GeForce RTX 5090 — Frequently Asked Questions
What is the NVIDIA GeForce RTX 5090 best used for?
When was the NVIDIA GeForce RTX 5090 released, and what was its launch MSRP?
Where do the benchmark numbers on this page come from?
Can the NVIDIA GeForce RTX 5090 run local LLMs?
Where can I buy the NVIDIA GeForce RTX 5090?
Buying guides that rank the NVIDIA GeForce RTX 5090's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the NVIDIA GeForce RTX 5090
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.
- RTX 3060 12GB Local LLM Guide: Which Models Actually Fit (2026)
- How to run Llama 3.1 8B on NVIDIA GeForce RTX 5090
- Gunnir Arc B580 vs RTX 5090D on DeepSeek: The Budget AI-Rig Upset Explained
- RTX 5090 vs RTX 5080: Is the $1000 Premium Worth It for 4K Gaming in 2026?
- RTX 5090 Prebuilt vs a $700 RTX 3060 Local-LLM Box: What Extra VRAM Actually Buys
- Tenstorrent TT-QuietBox 2 (Blackhole) vs RTX 5090: Should LLM Builders Care?
- Mistral Medium 3.5 Local Inference: Hardware Requirements and Benchmarks
- Ling 2.6 1T on Local Hardware: Can You Actually Run a Trillion-Parameter Model at Home in 2026?
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- How to Build a Windows 98 Retro PC in 2026
- Best 1440p Gaming GPUs in 2026
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- Best Retro Handhelds in 2026 — From $35 to $500
More reviews from the SpecPicks archive
Browse all reviews →- Set Up a Local LLM Coding Assistant in VS Code on an RTX 3060 12GB (2026)
- Putting a Modern SATA SSD in a Win98/XP Retro Build via an IDE Bridge
- Windows XP Gaming Laptops in 2026: A Realistic Buying Guide
- Best GPU for Running Llama 3 8B Locally Under $350 (2026)
- LiquidAI LFM2.5-8B-A1B: An 8B MoE You Can Run on a 12GB RTX 3060
- Ryzen 5 5600G vs Ryzen 7 5700X for Budget 1080p Gaming
- AI Bug-Hunters Are Flooding Security Reports: Running a Local Code-Audit LLM on an RTX 3060
- Best Controller for Forza Horizon 6 on PC (2026)
- RTX 3060 12GB for Local LLMs: The Complete 2026 Guide
- Self-Host Immich on a Raspberry Pi 4 8GB: 2026 Setup + Perf
- Best GPU for Local Stable Diffusion at 1080p/1440p in 2026
- Intel Arc B50 for Stable Diffusion: Setup & Performance
- Building a 2002 Windows XP Gaming Rig: GeForce 4 Ti, Sound BlasterX G6, CompactFlash Boot
- Best AMD Ryzen CPU for Gaming in 2026: 5 Picks Compared
- Is the RTX 3060 12GB Still a Good 1080p Gaming GPU in 2026?
- Logitech G29 vs HORI Force Feedback Wheel — Which First Wheel Is Right?
- Gemini 3.6 Flash Shipped: Why Local Builders Still Reach for a 12GB GPU
- Best GPU for Local Stable Diffusion Under $400: Why the RTX 3060 12GB Still Wins
- LEGO Batman: Legacy of the Dark Knight Steam Deck FPS
- GPT-5.6 Sol Reportedly Cracks a 30-Year Statistics Conjecture in 90 Minutes
- Can a Raspberry Pi 4 8GB Run a Local LLM with Ollama?
- Qwen3.7-Plus Goes Agentic: Cloud Model vs Your Local 12GB Rig
- Best SSD for Steam Deck OLED Expansion in 2026: SATA Adapter vs M.2 2230
- Claude Opus 4.8 Raised the Bar — Best Local Coding LLMs for a 12GB RTX 3060
More buying guides from SpecPicks
Browse all buying guides →- Best Mechanical Keyboards for Gaming in 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best GPUs for Running Local LLMs in 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best CPUs for Content Creators in 2026
- Best Gaming Mice for 2026
- Best NVMe External Enclosures for 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best PC Cases for Building in 2026
- Best CPU Coolers for 2026
- Best Gaming Monitors for 2026
- Best CPUs for Gaming in 2026
- Best Controllers for PC Gaming in 2026
- Best External SSDs for Content Creators in 2026
- Best AM5 Motherboards for 2026
- Best NVMe SSDs for Gaming in 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best GPUs for 4K Gaming in 2026
- Best Graphics Cards for Gaming in 2026
- Best 4K Monitors for Content Creators in 2026
- Best Retro Gaming Consoles & Handhelds for 2026