NVIDIA GeForce RTX 5070 — Benchmarks & Specs
*Price sourced from Amazon.com. Price and availability subject to change.
Bottom line: how fast is the NVIDIA GeForce RTX 5070?
At 1440p (Ultra, DLSS Quality), the NVIDIA GeForce RTX 5070 averages 204 fps in Cyberpunk 2077, per BabelTechReviews. For local LLM inference it generates 127.5 tokens/sec running llama2:7b at q4_0 under llama.cpp, per llama.cpp GitHub. In 3DMark Fire Strike it scores 41,611 points, per WorthPlaying. Its 12 GB of VRAM is the binding constraint for local inference: that capacity fits 8-13B-parameter models at Q4_K_M.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The NVIDIA GeForce RTX 5070 is a graphics card from the Blackwell family released in 2025 from NVIDIA. Key on-paper specs include 12 GB of GDDR7 VRAM, 250W TDP. It launched with a $549 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 11 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 5070, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).
Gaming Performance (measured FPS)
Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.
| Game | Resolution | Settings | Relative | Avg FPS | 1% low | Source |
|---|---|---|---|---|---|---|
| Final Fantasy XIV: Dawntrail | 1080p | Maximum | 225 fps | — | Gamers Nexus 2025-03-13 | |
| Resident Evil 4 | 1080p | Prioritize Graphics | 224 fps | — | Gamers Nexus 2025-03-13 | |
| Cyberpunk 2077 | 1440p | Ultra DLSS Quality | 204 fps | — | BabelTechReviews 2025-02-26 | |
| Final Fantasy XIV: Dawntrail | 1440p | Maximum | 172 fps | — | Gamers Nexus 2025-02-26 | |
| Alan Wake 2 | 1440p | Ultra DLSS Quality | 163 fps | — | BabelTechReviews 2025-02-26 | |
| Resident Evil 4 | 1440p | Prioritize Graphics | 152 fps | — | Gamers Nexus 2025-03-13 | |
| Resident Evil 4 | 1440p | Ultra Native | 149 fps | — | Gamers Nexus 2025-03-04 | |
| Resident Evil 4 | 1440p | Maximum RT RT on FSR Quality | 149 fps | — | Gamers Nexus 2025-03-13 | |
| DOOM Eternal | 4K | Ultra | 145 fps | — | BabelTech Reviews 2025-03-04 | |
| Cyberpunk 2077 | 1080p | Ultra | 138 fps | — | Gamers Nexus 2025-03-13 | |
| Cyberpunk 2077: Phantom Liberty | 1080p | Ultra | 138 fps | — | Gamers Nexus 2025-02-26 | |
| Cyberpunk 2077 | 4K | Ultra DLSS Quality | 133 fps | — | BabelTechReviews 2025-02-26 | |
| Cyberpunk 2077 | 4K | Ultra RT RT on DLSS Quality | 133 fps | — | BabelTech Reviews 2025-03-04 | |
| Hogwarts Legacy | 4K | Ultra DLSS Quality | 128 fps | — | BabelTech Reviews 2025-03-04 |
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| llama2:7b | q4_0 llama.cpp | 127.5 tok/s | — | llama.cpp GitHub 2025-03-01 | |
| llama2:7b | q4_0 llama.cpp | 127.5 tok/s | — | knightli.com 2026-04-23 | |
| llama3.2:1b | q4_K_M llama.cpp | 101.0 tok/s | — | LocalScore.ai 2025-04-01 | |
| Llama 3.2 1B Instruct | Q4_K_M llama.cpp | 101.0 tok/s | — | LocalScore 2025-06-01 | |
| qwen3:8b | q4_K_M llama.cpp | 59.1 tok/s | — | hardware-corner.net 2025-12-09 | |
| llama3.1:8b | q4_K_M llama.cpp | 55.9 tok/s | — | LocalScore.ai 2025-04-01 | |
| Llama 3.1 8B Instruct | Q4_K_M llama.cpp | 55.9 tok/s | — | LocalScore 2025-06-01 | |
| qwen3:0.6b | — ollama | 47.1 tok/s | — | LocalLLaMA 2026-04-15 | |
| qwen2.5:14b | q4_K_M llama.cpp | 20.8 tok/s | — | LocalScore.ai 2025-04-01 | |
| Qwen2.5 14B Instruct | Q4_K_M llama.cpp | 20.8 tok/s | — | LocalScore 2025-06-01 | |
| gemma:26b | q4_0 llama.cpp | 5.0 tok/s | — | LocalLLaMA 2026-04-16 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| 3DMark Fire Strike | 41,611 points | WorthPlaying 2025-03-04 | |
| PassMark G3D Mark | 28,758 pts | PassMark 2026-04-20 | |
| 3DMark Time Spy | 22,221 points | LAN OC 2025-03-04 | |
| 3DMark Time Spy | 22,221 points | LanOC Reviews 2025-03-04 | |
| 3DMark Port Royal | 13,939 points | WorthPlaying 2025-03-04 | |
| 3DMark Time Spy Extreme | 9,535 points | WorthPlaying 2025-03-04 | |
| 3DMark Speed Way | 5,805 points | LanOC Reviews 2025-02-26 | |
| 3DMark Speed Way | 5,769 points | WorthPlaying 2025-03-04 | |
| 3DMark Steel Nomad | 4,968 points | WorthPlaying 2025-03-04 | |
| PassMark G2D Mark | 1,299 pts | PassMark 2026-04-20 |
Products Featuring the NVIDIA GeForce RTX 5070
Full Specifications
| tdp w | 250 |
|---|---|
| vram gb | 12 |
| vram type | GDDR7 |
| cuda cores | 6144 |
NVIDIA GeForce RTX 5070 — Frequently Asked Questions
What is the NVIDIA GeForce RTX 5070 best used for?
When was the NVIDIA GeForce RTX 5070 released, and what was its launch MSRP?
Where do the benchmark numbers on this page come from?
Can the NVIDIA GeForce RTX 5070 run local LLMs?
Where can I buy the NVIDIA GeForce RTX 5070?
Buying guides that rank the NVIDIA GeForce RTX 5070's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the NVIDIA GeForce RTX 5070
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.
- Best 12GB GPU for Local LLMs in 2026
- Best GPU for Qwen 3 14B (2026)
- IBM Granite 4.1 8B vs Qwen 3.6 27B: Which Small Local Model Wins on a 16GB GPU?
- How to run DeepSeek-R1 32B on NVIDIA GeForce RTX 5070
- How to run Llama 3.1 70B on NVIDIA GeForce RTX 5070
- In Brief: Best Buy Cuts $1,000 Off RTX 5070 OLED Gaming Laptop with Core Ultra 9 275HX
- Best Mini-ITX Case for a Small Form Factor Gaming PC in 2026
- How to run Llama 3.1 8B on NVIDIA GeForce RTX 5070
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Best 1440p Gaming GPUs in 2026
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
- Best Retro Handhelds in 2026 — From $35 to $500
- How to Build a Windows 98 Retro PC in 2026
More reviews from the SpecPicks archive
Browse all reviews →- Ryzen AI Max Mini PC 128GB: Specs, Uses, Alternatives
- Self-Host Immich on a Budget Homelab Box: Storage Tiers That Make Sense
- KOORUI 27" 4K vs Samsung Odyssey 4K: Best 27-Inch Panel for PC and Console
- Gemini 3.5 Flash vs Local LLMs on a 12GB GPU: When Cloud Wins
- CompactFlash as a Silent IDE Boot Drive: A 1998-Era Build Log
- Intelligence Index v4.1: The Agentic-Benchmark Shift and Your Local Rig
- Grok 4.5 at $0.31 a Task: When Cheap Cloud Beats a Local Build
- A New Benchmark Says AI Fails at Real Knowledge Work — Does a Bigger Local Rig Help?
- Asus RTX 5060 Near-MSRP Deal Has Expired
- Jetson Orin Nano Super vs Raspberry Pi 5: Real Edge-AI Benchmarks (2026)
- Running DeepSeek Distills Locally on a Ryzen 7 5800X + RTX 3060
- Blue Yeti vs HyperX QuadCast 2 S: Best USB Mic for Streaming in 2026
- Best Cooler for the Ryzen 7 5800X: Noctua NH-U12S vs Cooler Master ML240L
- Raspberry Pi 5 vs Pi 4 8GB for Homelab: Which Should You Buy?
- Logitech Extreme 3D Pro vs DualSense for Flight and Space Sims in 2026
- Best SSD to Upgrade a PS4 Pro: Samsung 870 EVO vs Crucial BX500
- Best GPU for Llama 70B Local Inference in 2026: RTX 3060 12GB Dual vs RTX 3090 vs Gorgon Halo
- Best Raspberry Pi 5 Alternatives in 2025
- AutomationBench Cost Gap: What DeepSeek V4's 5-Cent Task Means for Local Agent Rigs
- Best Budget PC Gaming Upgrades Under $150 in 2026
- Ryzen AI Max+ 395 vs RTX 4070: Gaming vs AI in 2026
- RTX 3060 12GB vs Ryzen 5 5600G iGPU for Entry Local LLMs
- Old Massive Chieftec Full Tower: Homelab Gem or Junk?
- Linux Now Boots on a Sega Genesis — Here's What That Actually Means
More buying guides from SpecPicks
Browse all buying guides →- Best Gaming Monitors for 2026
- Best CPUs for Gaming in 2026
- Best External SSDs for Content Creators in 2026
- Best AM5 Motherboards for 2026
- Best Retro Gaming Consoles & Handhelds for 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best CPUs for Content Creators in 2026
- Best CPU Coolers for 2026
- Best NVMe External Enclosures for 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best PC Cases for Building in 2026
- Best GPUs for 4K Gaming in 2026
- Best Gaming Mice for 2026
- Best 4K Monitors for Content Creators in 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best GPUs for Running Local LLMs in 2026
- Best Graphics Cards for Gaming in 2026
- Best NVMe SSDs for Gaming in 2026
- Best Controllers for PC Gaming in 2026
- Best DDR5 RAM for Gaming PCs in 2026