NVIDIA GeForce RTX 3070 — Benchmarks & Specs
*Price sourced from Amazon.com. Price and availability subject to change.
Bottom line: how fast is the NVIDIA GeForce RTX 3070?
At 1440p (Ultra), the NVIDIA GeForce RTX 3070 averages 107 fps in Shadow of the Tomb Raider, per Tom's Hardware. For local LLM inference it generates 78.7 tokens/sec running llama2:7b at q4_0 under llama.cpp, per llama.cpp GitHub Discussion #10879. In PassMark G3D Mark it scores 22,123 pts, per PassMark. Its 8 GB of VRAM is the binding constraint for local inference: that capacity fits 7B-parameter models at Q4.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The NVIDIA GeForce RTX 3070 is a graphics card from the Ampere family released in 2020 from NVIDIA. Key on-paper specs include 8 GB of GDDR6 VRAM, 220W TDP. It launched with a $499 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 3070, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).
Gaming Performance (measured FPS)
Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.
| Game | Resolution | Settings | Relative | Avg FPS | 1% low | Source |
|---|---|---|---|---|---|---|
| The Witcher 3: Wild Hunt (Next-Gen) | 1080p | Ultra | 152 fps | — | AskGeek 2023-01-01 | |
| Cyberpunk 2077 | 1080p | Ultra | 107 fps | — | AskGeek 2023-10-01 | |
| Shadow of the Tomb Raider | 1440p | Ultra | 107 fps | — | Tom's Hardware 2020-10-29 | |
| Forza Horizon 5 | 1080p | Extreme | 97 fps | — | Notebookcheck 2021-11-14 | |
| Final Fantasy XIV: Dawntrail | 1440p | Ultra | 97 fps | — | Gamers Nexus 2025-04-17 | |
| Far Cry 6 | 1440p | Ultra | 94 fps | 73 fps | HardwareDB 2023-03-01 | |
| Cyberpunk 2077 | 1080p | Ultra | 85 fps | — | Gamers Nexus 2025-03-13 | |
| Forza Horizon 5 | 1440p | Extreme | 85 fps | — | Notebookcheck 2021-11-14 | |
| Dragon's Dogma 2 | 1080p | Ultra | 83 fps | — | Gamers Nexus 2025-04-17 | |
| Elden Ring | 1440p | Ultra | 82 fps | 60 fps | HardwareDB 2023-03-01 | |
| Crimson Desert | 1080p | Ultra | 81 fps | — | Gamers Nexus 2026-04-01 | |
| Baldur's Gate 3 | 1440p | Ultra | 80 fps | — | Gamers Nexus 2023-10-04 | |
| Assassin's Creed Valhalla | 1440p | Ultra | 79 fps | 54 fps | HardwareDB 2023-03-01 | |
| A Plague Tale: Requiem | 1440p | Ultra | 69 fps | 52 fps | HardwareDB 2023-03-01 |
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| llama2:7b | q4_0 llama.cpp | 78.7 tok/s | — | llama.cpp GitHub Discussion #10879 2024-12-18 | |
| llama2:7b | q4_0 llama.cpp | 78.7 tok/s | — | llama.cpp GitHub Discussions 2024-07-01 | |
| llama3.1:8b | q4_K_M ollama | 73.7 tok/s | 5.4 GB | Patrick Hughes (bmdpat.com) 2026-07-26 | |
| llama3:8b | q4_K_M llama.cpp | 72.8 tok/s | 4.6 GB | GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-07-01 | |
| llama3:8b | q4_K_M llama.cpp | 72.8 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub/XiongjieDai) 2024-05-01 | |
| llama3:8b | q4_K_M llama.cpp | 70.9 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub/XiongjieDai) 2024-05-01 | |
| llama3:8b | q4_K_M llama.cpp | 70.9 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-05-01 | |
| llama3:8b | q4_K_M llama.cpp | 67.0 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub/XiongjieDai) 2024-05-01 | |
| llama3:8b | q4_K_M llama.cpp | 67.0 tok/s | 4.6 GB | GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-07-01 | |
| llama3:8b | q4_K_M llama.cpp | 61.6 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub/XiongjieDai) 2024-05-01 | |
| llama3:8b | q4_K_M llama.cpp | 61.6 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-07-01 | |
| llama3.1:8b | q4_K_M llama.cpp | 59.6 tok/s | 8.0 GB | LocalScore.ai 2024-12-01 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| PassMark G3D Mark | 22,123 pts | PassMark 2026-04-20 | |
| 3DMark Fire Strike Extreme | 17,115 points | KitGuru 2020-10-23 | |
| 3DMark Fire Strike Extreme | 17,115 points | DigiStatement 2020-10-28 | |
| 3DMark Time Spy | 14,328 points | 3DMark 2021-06-01 | |
| 3DMark Time Spy | 14,048 points | KitGuru 2020-10-23 | |
| 3DMark Time Spy | 13,945 points | DigiStatement 2020-10-28 | |
| 3DMark Time Spy | 13,945 points | VideoCardz 2020-10-01 | |
| 3DMark Fire Strike Ultra | 8,749 points | Guru3D 2020-10-23 | |
| 3DMark Port Royal | 8,412 points | 3DMark 2022-04-11 | |
| 3DMark Port Royal | 8,324 points | DigiStatement 2020-10-28 |
Products Featuring the NVIDIA GeForce RTX 3070
Full Specifications
| tdp w | 220 |
|---|---|
| vram gb | 8 |
| vram type | GDDR6 |
| cuda cores | 5888 |
NVIDIA GeForce RTX 3070 — Frequently Asked Questions
What is the NVIDIA GeForce RTX 3070 best used for?
When was the NVIDIA GeForce RTX 3070 released, and what was its launch MSRP?
Where do the benchmark numbers on this page come from?
Can the NVIDIA GeForce RTX 3070 run local LLMs?
Where can I buy the NVIDIA GeForce RTX 3070?
Buying guides that rank the NVIDIA GeForce RTX 3070's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the NVIDIA GeForce RTX 3070
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.
- Arc B580 12GB vs RTX 3060 12GB: Which 12GB Card for 1440p in 2026?
- One Monitor, Two PCs: Monitor KVM vs Dual-Input for a Gaming Rig and AI Box
- The Greatest GPU Awards: A Decade of Standout Cards
- Alien: Isolation on Steam Deck: Best Settings for 2025
- GTX 1660 VRAM: Why 8GB Doesn't Exist (Real Specs)
- Best Wired Gear for Tournament-Legal LAN Play in 2026
- Logitech G29 vs HORI Racing Wheel Overdrive: Which Entry Sim Wheel Wins?
- How to Filter Steam Deck's 20K+ Games by FPS, Price, HLTB
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Best 1440p Gaming GPUs in 2026
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- Best Retro Handhelds in 2026 — From $35 to $500
- How to Build a Windows 98 Retro PC in 2026
More reviews from the SpecPicks archive
Browse all reviews →- GPT-5.5 Instant Shipped: What an RTX 3060 12GB Local Stack Covers When OpenAI Retires a Model
- GeForce 6800 Ultra AGP Install Guide: Drivers, Benchmarks, and the Last Great AGP Card
- Build a 2001 GeForce 3 Windows 98 Gaming PC in 2026: Parts and Setup
- Best GPU for 1440p Esports in 2026: Why the RTX 3060 12GB Still Delivers
- RTX 5070 Ti vs 4070 Ti SUPER
- GeForce4 Ti 4600 tuning guide — ForceWare, Detonator, and XP/98 driver pairing
- RTX 5070 Ti vs RTX 5080: Is the $400 Step-Up Worth It at 1440p and 4K?
- Jellyfin vs Plex on a Raspberry Pi 4 (8GB) in 2026: Transcoding, Power Draw, and Which to Self-Host
- Started a Homelab a Month Ago — Is a Ryzen 5 5600G Enough?
- Crucial BX500 vs Samsung 970 EVO Plus: SATA vs NVMe Boot Drive
- Microsoft + Nvidia AI PCs Run Real Agents: The Local Hardware That Matches (2026)
- AI-Driven Driver Recovery for SB Live! and Audigy on Win98: How an LLM Watches the Installer
- Ideogram 4.0 Open Weights on an RTX 3060 12GB: Local Text-to-Image in 2026
- Can a Raspberry Pi 4 8GB Run a Local LLM in 2026? Realistic tok/s
- CompactFlash + IDE Storage for a Period-Correct Windows 98 Build
- RX 9070 XT vs RTX 3060 12GB for Local LLM Inference (2026)
- Best Storage Upgrades for Retro and Budget PC Builds in 2026
- GameSir G7 SE vs 8BitDo Pro 2: Best Wired-Feel Controller for PC & Steam Deck
- Build a Retro Emulation Cabinet Brain on a Pi 4 8GB in 2026
- HyperX QuadCast 2 S vs Blue Yeti for Streamers: Which USB Mic Wins in 2026?
- Best Gaming GPUs for 1080p High-Refresh Rate (2026)
- Best Budget PC Upgrades Under $250 in 2026
- Glide vs OpenGL vs Direct3D: The PC Gaming API War 1996-2003, Benchmarked
- Ollama vs llama.cpp for Qwen 3.6 27B on a 12GB RTX 3060
More buying guides from SpecPicks
Browse all buying guides →- Best External SSDs for Content Creators in 2026
- Best Graphics Cards for Gaming in 2026
- Best Gaming Monitors for 2026
- Best NVMe SSDs for Gaming in 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best CPUs for Content Creators in 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best 4K Monitors for Content Creators in 2026
- Best GPUs for 4K Gaming in 2026
- Best Gaming Mice for 2026
- Best GPUs for Running Local LLMs in 2026
- Best AM5 Motherboards for 2026
- Best CPUs for Gaming in 2026
- Best CPU Coolers for 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best PC Cases for Building in 2026
- Best NVMe External Enclosures for 2026
- Best Controllers for PC Gaming in 2026
- Best Retro Gaming Consoles & Handhelds for 2026