Skip to main content
NVIDIA GeForce RTX 5070 Ti
NVIDIA · GPU · Blackwell

NVIDIA GeForce RTX 5070 Ti — Benchmarks & Specs

16 GB VRAM300W TDP$749 MSRP2025

*Price sourced from Amazon.com. Price and availability subject to change.

Bottom line: how fast is the NVIDIA GeForce RTX 5070 Ti?

At 1440p (Ultra DLSS Balanced + MFG 4x, DLSS Balanced), the NVIDIA GeForce RTX 5070 Ti averages 424 fps in Cyberpunk 2077, per Digital Trends. For local LLM inference it generates 189.4 tokens/sec running meta-llama/gpt-oss-20b-MoE at MXFP4 under llama.cpp, per llama.cpp GitHub Discussions. In PassMark G3D Mark it scores 32,422 pts, per PassMark. Its 16 GB of VRAM is the binding constraint for local inference: that capacity fits 13-17B-parameter models at Q4 with room for long context.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The NVIDIA GeForce RTX 5070 Ti is a graphics card from the Blackwell family released in 2025 from NVIDIA. Key on-paper specs include 16 GB of GDDR7 VRAM, 300W TDP. It launched with a $749 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 5070 Ti, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the NVIDIA GeForce RTX 5070 Ti by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Cyberpunk 2077 1440p Ultra DLSS Balanced + MFG 4x DLSS Balanced 424 fps Digital Trends 2025-03-24
War Thunder 1440p Ultra Native 378 fps TechSpot 2025-02-19
Alan Wake 2 1440p RT DLSS Balanced + MFG 4x RT on DLSS Balanced 340 fps Digital Trends 2025-03-24
Doom Eternal 1440p Ultra Nightmare RT on 332 fps 217 fps PCGamesN 2025-02-19
Doom: The Dark Ages 1440p Ultra Nightmare RT on DLSS Quality 285 fps PCGamesN 2025-02-19
Final Fantasy XIV 4K Maximum DLSS + MFG 4x DLSS 280 fps LanOC 2025-02-19
Cyberpunk 2077 4K Ultra DLSS Balanced + MFG 4x DLSS Balanced 250 fps Digital Trends 2025-03-24
Final Fantasy XIV 1440p Maximum Native 187 fps Gamers Nexus 2025-02-21
Final Fantasy XIV: Dawntrail 1440p Maximum 187 fps Gamers Nexus 2025-02-19
Final Fantasy XIV 1440p Maximum 187 fps Gamers Nexus 2025-02-21
Doom Eternal 4K Ultra Nightmare RT on 184 fps PCGamesN 2025-02-19
Final Fantasy XIV 4K Maximum DLSS + MFG 2x DLSS 172 fps LanOC 2025-02-19
Cyberpunk 2077 1440p Ultra DLSS Balanced DLSS Balanced 152 fps Digital Trends 2025-03-24
Indiana Jones and the Great Circle 1080p Ultra Native 148 fps PCGamesN 2025-02-20

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the NVIDIA GeForce RTX 5070 Ti, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
meta-llama/gpt-oss-20b-MoE MXFP4 llama.cpp 189.4 tok/s llama.cpp GitHub Discussions 2025-08-01
llama3.2:1b q4_K_M llama.cpp 185.0 tok/s LocalScore.ai 2025-03-01
meta-llama/Llama-3.2-1B-Instruct q4_K_M llama.cpp 185.0 tok/s LocalScore.ai 2025-04-01
llama2:7b q4_0 llama.cpp 182.4 tok/s KnightLi 2026-04-23
llama3.1:8b q4_K_M ollama 125.0 tok/s 4.9 GB ComputingForGeeks 2026-06-28
qwen3:8b q4_K_M llama.cpp 120.5 tok/s Hardware Corner 2026-03-01
mistral:7b q4_K_M ollama 112.0 tok/s 4.8 GB Compute Market 2026-03-28
llama3.1:8b q8_0 ollama 78.0 tok/s 8.5 GB Compute Market 2026-03-28
meta-llama/Llama-3.1-8B-Instruct q5_K_M ollama 75.0 tok/s 8.0 GB ModelFit.io 2025-04-01
qwen3:14b q4_K_M llama.cpp 74.3 tok/s Hardware Corner 2026-03-01
qwen2.5:14b q4_K_M ollama 73.2 tok/s 9.0 GB ComputingForGeeks 2026-06-28
llama3.1:8b q4_K_M llama.cpp 65.5 tok/s LocalScore.ai 2025-03-01

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the NVIDIA GeForce RTX 5070 Ti — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
PassMark G3D Mark 32,422 pts PassMark 2026-04-20
3DMark Time Spy 27,949 points LanOC 2025-02-19
3DMark Time Spy 27,949 points LanOC Reviews 2025-02-19
3DMark Time Spy 25,610 points Digital Trends 2025-02-19
3DMark Port Royal 18,851 points HWBot 2025-05-27
3DMark Port Royal 17,461 points Digital Trends 2025-02-19
3DMark Speed Way 7,708 points Overclocking.com 2025-02-22
3DMark Steel Nomad 6,932 points Overclocking.com 2025-04-18
3DMark Steel Nomad 6,905 points UL Benchmarks 2025-03-01
PassMark G2D Mark 1,329 pts PassMark 2026-04-20

Products Featuring the NVIDIA GeForce RTX 5070 Ti

Full Specifications

tdp w300
vram gb16
vram typeGDDR7
cuda cores8960

NVIDIA GeForce RTX 5070 Ti — Frequently Asked Questions

What is the NVIDIA GeForce RTX 5070 Ti best used for?
NVIDIA GeForce RTX 5070 Ti is positioned as a 16 GB VRAM Blackwell-family graphics card. Use it for 1440p/4K gaming and mid-tier AI workloads. See the synthetic + AI benchmark tables below for measured performance.
When was the NVIDIA GeForce RTX 5070 Ti released, and what was its launch MSRP?
NVIDIA GeForce RTX 5070 Ti launched in 2025 at a $749 MSRP. Street prices diverge from launch pricing over a product's lifetime — check the linked Amazon listings on this page for current availability.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the NVIDIA GeForce RTX 5070 Ti run local LLMs?
Yes — NVIDIA GeForce RTX 5070 Ti has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). 16 GB VRAM is enough for 13-17B parameter models at Q4 with comfortable context.
Where can I buy the NVIDIA GeForce RTX 5070 Ti?
Amazon listings are linked at the top of this page (in the hero CTA) and in the "Products Featuring this Hardware" section below. SpecPicks earns a small affiliate commission on qualifying purchases.

Buying guides that rank the NVIDIA GeForce RTX 5070 Ti's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the NVIDIA GeForce RTX 5070 Ti

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →
NVIDIA GeForce RTX 5070 Ti
NVIDIA GeForce RTX 5070 Ti
$1698.88
View on Amazon →