Skip to main content
NVIDIA GeForce RTX 5070
NVIDIA · GPU · Blackwell

NVIDIA GeForce RTX 5070 — Benchmarks & Specs

12 GB VRAM250W TDP$549 MSRP2025

*Price sourced from Amazon.com. Price and availability subject to change.

Bottom line: how fast is the NVIDIA GeForce RTX 5070?

At 1440p (Ultra, DLSS Quality), the NVIDIA GeForce RTX 5070 averages 204 fps in Cyberpunk 2077, per BabelTechReviews. For local LLM inference it generates 127.5 tokens/sec running llama2:7b at q4_0 under llama.cpp, per llama.cpp GitHub. In 3DMark Fire Strike it scores 41,611 points, per WorthPlaying. Its 12 GB of VRAM is the binding constraint for local inference: that capacity fits 8-13B-parameter models at Q4_K_M.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The NVIDIA GeForce RTX 5070 is a graphics card from the Blackwell family released in 2025 from NVIDIA. Key on-paper specs include 12 GB of GDDR7 VRAM, 250W TDP. It launched with a $549 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 11 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 5070, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the NVIDIA GeForce RTX 5070 by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Final Fantasy XIV: Dawntrail 1080p Maximum 225 fps Gamers Nexus 2025-03-13
Resident Evil 4 1080p Prioritize Graphics 224 fps Gamers Nexus 2025-03-13
Cyberpunk 2077 1440p Ultra DLSS Quality 204 fps BabelTechReviews 2025-02-26
Final Fantasy XIV: Dawntrail 1440p Maximum 172 fps Gamers Nexus 2025-02-26
Alan Wake 2 1440p Ultra DLSS Quality 163 fps BabelTechReviews 2025-02-26
Resident Evil 4 1440p Prioritize Graphics 152 fps Gamers Nexus 2025-03-13
Resident Evil 4 1440p Ultra Native 149 fps Gamers Nexus 2025-03-04
Resident Evil 4 1440p Maximum RT RT on FSR Quality 149 fps Gamers Nexus 2025-03-13
DOOM Eternal 4K Ultra 145 fps BabelTech Reviews 2025-03-04
Cyberpunk 2077 1080p Ultra 138 fps Gamers Nexus 2025-03-13
Cyberpunk 2077: Phantom Liberty 1080p Ultra 138 fps Gamers Nexus 2025-02-26
Cyberpunk 2077 4K Ultra DLSS Quality 133 fps BabelTechReviews 2025-02-26
Cyberpunk 2077 4K Ultra RT RT on DLSS Quality 133 fps BabelTech Reviews 2025-03-04
Hogwarts Legacy 4K Ultra DLSS Quality 128 fps BabelTech Reviews 2025-03-04

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the NVIDIA GeForce RTX 5070, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
llama2:7b q4_0 llama.cpp 127.5 tok/s llama.cpp GitHub 2025-03-01
llama2:7b q4_0 llama.cpp 127.5 tok/s knightli.com 2026-04-23
llama3.2:1b q4_K_M llama.cpp 101.0 tok/s LocalScore.ai 2025-04-01
Llama 3.2 1B Instruct Q4_K_M llama.cpp 101.0 tok/s LocalScore 2025-06-01
qwen3:8b q4_K_M llama.cpp 59.1 tok/s hardware-corner.net 2025-12-09
llama3.1:8b q4_K_M llama.cpp 55.9 tok/s LocalScore.ai 2025-04-01
Llama 3.1 8B Instruct Q4_K_M llama.cpp 55.9 tok/s LocalScore 2025-06-01
qwen3:0.6b ollama 47.1 tok/s LocalLLaMA 2026-04-15
qwen2.5:14b q4_K_M llama.cpp 20.8 tok/s LocalScore.ai 2025-04-01
Qwen2.5 14B Instruct Q4_K_M llama.cpp 20.8 tok/s LocalScore 2025-06-01
gemma:26b q4_0 llama.cpp 5.0 tok/s LocalLLaMA 2026-04-16

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the NVIDIA GeForce RTX 5070 — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
3DMark Fire Strike 41,611 points WorthPlaying 2025-03-04
PassMark G3D Mark 28,758 pts PassMark 2026-04-20
3DMark Time Spy 22,221 points LAN OC 2025-03-04
3DMark Time Spy 22,221 points LanOC Reviews 2025-03-04
3DMark Port Royal 13,939 points WorthPlaying 2025-03-04
3DMark Time Spy Extreme 9,535 points WorthPlaying 2025-03-04
3DMark Speed Way 5,805 points LanOC Reviews 2025-02-26
3DMark Speed Way 5,769 points WorthPlaying 2025-03-04
3DMark Steel Nomad 4,968 points WorthPlaying 2025-03-04
PassMark G2D Mark 1,299 pts PassMark 2026-04-20

Products Featuring the NVIDIA GeForce RTX 5070

Full Specifications

tdp w250
vram gb12
vram typeGDDR7
cuda cores6144

NVIDIA GeForce RTX 5070 — Frequently Asked Questions

What is the NVIDIA GeForce RTX 5070 best used for?
NVIDIA GeForce RTX 5070 is positioned as a 12 GB VRAM Blackwell-family graphics card. Use it for 1080p/1440p gaming and lightweight content creation. See the synthetic + AI benchmark tables below for measured performance.
When was the NVIDIA GeForce RTX 5070 released, and what was its launch MSRP?
NVIDIA GeForce RTX 5070 launched in 2025 at a $549 MSRP. Street prices diverge from launch pricing over a product's lifetime — check the linked Amazon listings on this page for current availability.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the NVIDIA GeForce RTX 5070 run local LLMs?
Yes — NVIDIA GeForce RTX 5070 has 11 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). 12 GB VRAM is well-suited to 8-13B parameter models at Q4_K_M.
Where can I buy the NVIDIA GeForce RTX 5070?
Amazon listings are linked at the top of this page (in the hero CTA) and in the "Products Featuring this Hardware" section below. SpecPicks earns a small affiliate commission on qualifying purchases.

Buying guides that rank the NVIDIA GeForce RTX 5070's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the NVIDIA GeForce RTX 5070

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →
NVIDIA GeForce RTX 5070
NVIDIA GeForce RTX 5070
$1698.88
View on Amazon →