GeForce RTX 4080 — Benchmarks & Specs
*Price sourced from Amazon.com. Price and availability subject to change.
Bottom line: how fast is the GeForce RTX 4080?
At 1440p (Ultra), the GeForce RTX 4080 averages 255 fps in Shadow of the Tomb Raider, per Hardware Times. For local LLM inference it generates 147.5 tokens/sec running qwen3:30b-a3b at IQ3_XXS under llama.cpp, per Glukhov.org — 16GB VRAM LLM benchmarks (llama.cpp). In 3DMark Fire Strike it scores 48,524 points, per Overclocking.com. Its 16 GB of VRAM is the binding constraint for local inference: that capacity fits 13-17B-parameter models at Q4 with room for long context.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The GeForce RTX 4080 is a graphics card from the Ada Lovelace family from NVIDIA. Key on-paper specs include 16 GB of GDDR6X VRAM, 320W TDP. Data on this page draws on 8+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the GeForce RTX 4080, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).
Gaming Performance (measured FPS)
Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.
| Game | Resolution | Settings | Relative | Avg FPS | 1% low | Source |
|---|---|---|---|---|---|---|
| Shadow of the Tomb Raider | 1440p | Ultra | 255 fps | — | Hardware Times 2024-03-19 | |
| Forza Horizon 5 | 1080p | Ultra | 241 fps | 179 fps | DropReference 2026-08-01 | |
| F1 2022 | 1440p | Ultra | 232 fps | — | Hardware Times 2024-03-19 | |
| Hitman 3 | 1440p | Ultra | 205 fps | — | Hardware Times 2024-03-19 | |
| Horizon Zero Dawn | 1440p | Ultimate | 200 fps | — | KitGuru 2022-11-16 | |
| Ghostwire: Tokyo | 1440p | Ultra | 193 fps | — | Hardware Times 2024-03-19 | |
| Tiny Tina's Wonderlands | 1440p | Ultra | 183 fps | — | Hardware Times 2024-03-19 | |
| Shadow of the Tomb Raider | 1440p | RT Ultra RT on | 180 fps | — | Hardware Times 2024-03-19 | |
| Dying Light 2 | 1440p | Ultra | 176 fps | — | Hardware Times 2024-03-19 | |
| Shadow of the Tomb Raider | 4K | Ultra RT on DLSS Performance | 174 fps | — | Overclocking.com 2022-11-16 | |
| Dying Light 2 | 1440p | High | 170 fps | — | KitGuru 2022-11-16 | |
| Cyberpunk 2077 | 1440p | RT Psycho RT on DLSS Quality | 168 fps | — | Hardware Times 2024-03-19 | |
| Assassin's Creed Mirage | 1080p | Ultra | 165 fps | 122 fps | Hardware Times 2023-10-05 | |
| F1 2022 | 4K | Ultra | 164 fps | — | Hardware Times 2024-03-19 |
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| qwen3:30b-a3b | IQ3_XXS llama.cpp | 147.5 tok/s | 13.8 GB | Glukhov.org — 16GB VRAM LLM benchmarks (llama.cpp) 2026-03-09 | |
| llama2:7b | Q4_0 llama.cpp | 143.5 tok/s | — | knightli.com 2026-04-23 | |
| llama2:7b | q4_0 llama.cpp | 142.5 tok/s | — | llama.cpp GitHub 2025-08-07 | |
| llama2:7b | q4_0 llama.cpp | 142.5 tok/s | — | llama.cpp GitHub Discussion #15013 2025-08-07 | |
| gpt-oss:20b | q4_K_M ollama | 139.9 tok/s | 14.0 GB | glukhov.org 2026-03-09 | |
| gpt-oss:20b | Q4_K_M ollama | 139.9 tok/s | 14.0 GB | Rost Glukhov 2026-03-09 | |
| gpt-oss:20b | MXFP4 llama.cpp | 136.5 tok/s | 14.0 GB | Hardware Corner 2025-09-01 | |
| gemma2:27b | IQ4_XS llama.cpp | 121.7 tok/s | 14.7 GB | Glukhov.org — 16GB VRAM LLM benchmarks (llama.cpp) 2026-03-09 | |
| llama3.1:8b | Q4_K_M ollama | 117.0 tok/s | 4.9 GB | Markaicode 2024-06-01 | |
| llama3:8b | q4_K_M llama.cpp | 106.2 tok/s | — | GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-05-01 | |
| qwen3:8b | q4_K_M llama.cpp | 102.7 tok/s | — | Hardware Corner 2025-09-01 | |
| qwen3:8b | Q4_K_M llama.cpp | 102.7 tok/s | — | Hardware Corner 2025-06-01 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| 3DMark Fire Strike | 48,524 points | Overclocking.com 2022-11-15 | |
| PassMark G3D Mark | 34,433 pts | PassMark 2026-04-20 | |
| 3DMark Time Spy | 28,599 points | Tom's Hardware 2022-11-03 | |
| 3DMark Time Spy | 26,185 points | 3DMark 2023-01-01 | |
| 3DMark Port Royal | 17,663 points | 3DMark 2022-12-25 | |
| 3DMark Port Royal | 17,650 points | WCCFTech 2022-11-01 | |
| 3DMark Port Royal | 17,600 points | HyperCyber 2022-11-03 | |
| 3DMark Time Spy Extreme | 14,178 points | WCCFTech 2022-11-01 | |
| 3DMark Time Spy Extreme | 13,977 points | Wccftech 2022-10-05 | |
| 3DMark Speed Way | 7,532 points | 3DMark 2023-12-25 |
Products Featuring the GeForce RTX 4080
Full Specifications
| rops | 112 |
|---|---|
| tmus | 304 |
| tdp w | 320 |
| opengl | 4.6 |
| vulkan | 1.4 |
| directx | 12 Ultimate (12_2) |
| foundry | TSMC |
| outputs | 1x HDMI 2.1 3x DisplayPort 1.4a |
| vram gb | 16 |
| gpu chip | AD103 |
| l2 cache | 64 MB |
| rt cores | 76 |
| sm count | 76 |
| bandwidth | 716.8 GB/s |
| vram type | GDDR6X |
| cuda cores | 9728 |
| generation | GeForce 40 |
| process nm | 5 |
| gpu variant | AD103-300-A1 |
| spec source | TechPowerUp GPU Database |
GeForce RTX 4080 — Frequently Asked Questions
What is the GeForce RTX 4080 best used for?
When was the GeForce RTX 4080 released, and what was its launch MSRP?
Where do the benchmark numbers on this page come from?
Can the GeForce RTX 4080 run local LLMs?
Where can I buy the GeForce RTX 4080?
Buying guides that rank the GeForce RTX 4080's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the GeForce RTX 4080
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.
- Real Productivity on 32-64GB RAM for Local LLMs
- RTX 3090 vs RTX 4090 for LLM Inference: Same 24GB (2026)
- Best GPU for Gemma 3 27B in 2026: Where 12 GB Stops Being Enough
- Qwen 0.8B AI Detectors: Fine-Tuned on Pangram-Style Data
- OpenAI Names Its Biggest Data Center Yet, With NVIDIA Backing: What It Means
- The Pac-Man Benchmark: Testing Local Agentic Coding AI
- llama.cpp vs vLLM for Single-User Local Chat in 2026: Which Backend Fits?
- Why Privacy-Preserving AI Demand Is Rising in the LLM Era
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- Best Retro Handhelds in 2026 — From $35 to $500
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- Best 1440p Gaming GPUs in 2026
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
- How to Build a Windows 98 Retro PC in 2026
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
More reviews from the SpecPicks archive
Browse all reviews →- Ryzen 7 5800X vs 5700X for a Dual-Duty Gaming and Local-LLM Build
- Best GPU for Llama 3.1 405B (2026)
- Add AI Vision to a Raspberry Pi 4 8GB with an AI Accelerator (2026)
- Top 20 Most Played Steam Deck Games: April 2026
- Using an LLM to Fix Win98 Voodoo & TNT Driver Installs
- Best DualSense and PC-Compatible Controller (2026)
- Build a Privacy-First Ring Alternative on a Raspberry Pi 4 in 2026
- Best GPU for DeepSeek-R1 32B (2026)
- DeepSeek on the US Entity List: What It Means for Local Inference
- Forza Horizon 6: 8GB vs 16GB VRAM GPU Benchmark Guide
- Qwen3.6-27B on Dual RTX 3060 12GB: The $400 30-50 tok/s Local LLM Build
- Logitech G920 vs HORI Racing Wheel: Best First Sim Wheel?
- Adafruit Reverse-Engineers the Creative Katana V2X Soundbar
- Fix Windows 98 SE on 1GB+ RAM: The Vcache Gotcha
- Best Gaming Keyboard for Office and Gaming Crossover (2026)
- Stable Diffusion WebUI Forge on an RTX 3060 12GB: Setup and Real Throughput
- First PC Build Guide: Components, Budget & Aesthetics
- Samsung Unveils World's First 360Hz 4K QD-OLED Gaming Panel
- GLM-5.2 vs DeepSeek V4 on a 12GB RTX 3060: Which Open-Weights Model Wins?
- Amazon RTX 5060 Deal: $324 Budget GPU for 1080p Gaming
- Claude Opus 4.8 Tops GPT-5.5: What Runs Local on a 12GB GPU
- Sony Trinitron FW900 Setup for Retro Gaming: Resolution Sweet Spots, Calibration, and Refresh-Rate Math
- 8BitDo Pro 2 vs DualSense for PC Emulation: Which Controller Wins
- Dual RTX 3090 vs RTX 5090: Gaming vs AI Training
More buying guides from SpecPicks
Browse all buying guides →- Best GPUs for 4K Gaming in 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best Gaming Monitors for 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best PC Cases for Building in 2026
- Best Graphics Cards for Gaming in 2026
- Best 4K Monitors for Content Creators in 2026
- Best NVMe External Enclosures for 2026
- Best NVMe SSDs for Gaming in 2026
- Best CPU Coolers for 2026
- Best CPUs for Gaming in 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best GPUs for Running Local LLMs in 2026
- Best External SSDs for Content Creators in 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best Gaming Mice for 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best Controllers for PC Gaming in 2026
- Best Retro Gaming Consoles & Handhelds for 2026
- Best CPUs for Content Creators in 2026
- Best AM5 Motherboards for 2026