NVIDIA GeForce RTX 4080 SUPER — Benchmarks & Specs
*Price sourced from Amazon.com. Price and availability subject to change.
Bottom line: how fast is the NVIDIA GeForce RTX 4080 SUPER?
At 1440p (Ultra, DLSS Quality), the NVIDIA GeForce RTX 4080 SUPER averages 172 fps in Avatar: Frontiers of Pandora with a 138 fps 1% low, per WorthPlaying. For local LLM inference it generates 147.1 tokens/sec running llama2:7b at q4_0 under llama.cpp, per llama.cpp GitHub Discussions. In 3DMark Fire Strike it scores 55,495 points, per LanOC Reviews. Its 16 GB of VRAM is the binding constraint for local inference: that capacity fits 13-17B-parameter models at Q4 with room for long context.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The NVIDIA GeForce RTX 4080 SUPER is a graphics card from the Ada Lovelace family released in 2024 from NVIDIA. Key on-paper specs include 16 GB of GDDR6X VRAM, 320W TDP. It launched with a $999 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 8+ Amazon listings, 9 synthetic benchmark results, 11 community AI inference reports, 14 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA GeForce RTX 4080 SUPER, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).
Gaming Performance (measured FPS)
Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.
| Game | Resolution | Settings | Relative | Avg FPS | 1% low | Source |
|---|---|---|---|---|---|---|
| Avatar: Frontiers of Pandora | 1440p | Ultra DLSS Quality | 172 fps | 138 fps | WorthPlaying 2024-01-31 | |
| Call of Duty: Modern Warfare III | 1440p | Ultra | 149 fps | 126 fps | WorthPlaying 2024-01-31 | |
| Avatar: Frontiers of Pandora | 1440p | RT Ultra RT on DLSS Quality | 148 fps | 117 fps | WorthPlaying 2024-01-31 | |
| The Callisto Protocol | 1440p | Ultra | 142 fps | 60 fps | WorthPlaying 2024-01-31 | |
| Far Cry 6 | 4K | Ultra RT on FSR Balanced | 134 fps | — | Overclocking.com 2024-01-31 | |
| Cyberpunk 2077 | 1440p | Ultra DLSS Quality | 132 fps | 66 fps | WorthPlaying 2024-01-31 | |
| Avatar: Frontiers of Pandora | 1440p | Ultra | 124 fps | 98 fps | WorthPlaying 2024-01-31 | |
| Cyberpunk 2077 | 1440p | Ultra | 115 fps | 76 fps | WorthPlaying 2024-01-31 | |
| Marvel's Spider-Man Remastered | 4K | Very High RT on DLSS Quality | 111 fps | — | KitGuru 2024-01-31 | |
| Avatar: Frontiers of Pandora | 4K | Ultra DLSS Quality | 108 fps | 89 fps | WorthPlaying 2024-01-31 | |
| Cyberpunk 2077 | 4K | RT Ultra RT on DLSS Balanced | 104 fps | — | Overclocking.com 2024-01-31 | |
| Red Dead Redemption II | 1440p | Ultra | 104 fps | 38 fps | WorthPlaying 2024-01-31 | |
| Call of Duty: Modern Warfare III | 4K | Ultra | 104 fps | 85 fps | WorthPlaying 2024-01-31 | |
| Forza Motorsport | 4K | Ultra | 100 fps | — | Hardware Times 2024-04-26 |
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| llama2:7b | q4_0 llama.cpp | 147.1 tok/s | — | llama.cpp GitHub Discussions 2024-12-18 | |
| qwen3:8b | q4_K_M llama.cpp | 102.7 tok/s | — | Hardware Corner 2025-06-01 | |
| llama3.1:8b | q4_K_M llama.cpp | 102.0 tok/s | 5.5 GB | MyAIHardware 2026-05-01 | |
| qwen3:8b | q4_K_M llama.cpp | 77.9 tok/s | — | Hardware Corner 2025-06-01 | |
| qwen3:14b | q4_K_M llama.cpp | 62.0 tok/s | — | Hardware Corner 2025-06-01 | |
| qwen2.5:14b | q4_K_M llama.cpp | 60.0 tok/s | — | kunalganglani.com LLM Benchmarks 2024-09-01 | |
| llama3.1:8b | q4_K_M llama.cpp | 54.4 tok/s | — | LocalScore.ai 2024-06-01 | |
| qwen2.5:14b | q4_K_M llama.cpp | 43.6 tok/s | — | LocalScore.ai 2024-06-01 | |
| qwen3:97b | Q5 llama.cpp | 29.0 tok/s | — | LocalLLaMA 2026-04-13 | |
| qwen3:235b | Q3 llama.cpp | 11.0 tok/s | — | LocalLLaMA 2026-03-28 | |
| gemma:26b | q4_0 llama.cpp | 5.0 tok/s | — | LocalLLaMA 2026-04-16 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| 3DMark Fire Strike | 55,495 points | LanOC Reviews 2024-01-31 | |
| PassMark G3D Mark | 34,256 pts | PassMark 2026-04-20 | |
| 3DMark Time Spy | 27,024 points | Overclocking.com 2024-01-31 | |
| 3DMark Port Royal | 18,593 points | Overclocking.com 2024-01-31 | |
| 3DMark Time Spy Extreme | 13,324 points | Overclocking.com 2024-01-31 | |
| 3DMark Speed Way | 7,558 points | Overclocking.com 2024-01-31 | |
| 3DMark Steel Nomad | 6,621 points | UL Benchmarks (3DMark) 2024-06-01 | |
| PassMark G2D Mark | 1,279 pts | PassMark 2026-04-20 | |
| Tom's Hardware GPU Hierarchy | 1st place | Tom's Hardware 2026-04-20 |
Products Featuring the NVIDIA GeForce RTX 4080 SUPER
Full Specifications
| tdp w | 320 |
|---|---|
| vram gb | 16 |
| vram type | GDDR6X |
| cuda cores | 10240 |
NVIDIA GeForce RTX 4080 SUPER — Frequently Asked Questions
What is the NVIDIA GeForce RTX 4080 SUPER best used for?
When was the NVIDIA GeForce RTX 4080 SUPER released, and what was its launch MSRP?
Where do the benchmark numbers on this page come from?
Can the NVIDIA GeForce RTX 4080 SUPER run local LLMs?
Where can I buy the NVIDIA GeForce RTX 4080 SUPER?
Buying guides that rank the NVIDIA GeForce RTX 4080 SUPER's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the NVIDIA GeForce RTX 4080 SUPER
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.
- RTX 5080 vs RTX 4080 SUPER — one generation apart
- Best 16GB GPU for Local LLM 2026
- Real Productivity on 32-64GB RAM for Local LLMs
- Qwen 0.8B AI Detectors: Fine-Tuned on Pangram-Style Data
- Best GPU for Gemma 3 27B in 2026: Where 12 GB Stops Being Enough
- RTX 3090 vs RTX 4090 for LLM Inference: Same 24GB (2026)
- The Pac-Man Benchmark: Testing Local Agentic Coding AI
- OpenAI Codex 'Watch Once, Repeat Forever': What It Means for Local Coding Rigs
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- Best 1440p Gaming GPUs in 2026
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
- How to Build a Windows 98 Retro PC in 2026
- Best Retro Handhelds in 2026 — From $35 to $500
More reviews from the SpecPicks archive
Browse all reviews →- Qwen3.6-27B at 80 TPS on RTX 5090: Is the Claim Real?
- 180Hz 1440p Monitors Hit Entry Pricing — Here's How the 4K Tier Compares
- Running Qwen3.6 35B A3B at 80 tok/s on a 12GB GPU: What the MSI RTX 3060 12GB Setup Looks Like
- Best Budget 4K Monitor in 2026: SANSUI vs KOORUI vs ASUS TUF
- AVX-512 Speeds Linux Software RAID Up to 41% — What It Means for Your Homelab NAS
- Building a 1999 GeForce 256 + Pentium III Win98 Rig in 2026
- Downsized Your Homelab? Here's What It's Worth
- Best Wireless Gaming Controllers for PC and Console in 2026
- First Homelab Setup: Hardware, Security & Networking Guide
- Best Budget SSD for Gaming and Boot Drives in 2026
- Blue Yeti vs HyperX QuadCast 2 S: Best Streaming Mic Under $200
- Best Parts for a DIY Home Arcade Cabinet in 2026
- CPU Fan Angled Left on a New Prebuilt: Normal or Not?
- Best GPU for Stable Diffusion Under $400: Why the RTX 3060 12GB Still Wins
- Panther Lake NPU vs RTX 3060: Which Runs Local LLMs Faster?
- Best Way to Stream a PS4 Pro to PC: Elgato Cam Link 4K Capture Setup
- How to run Qwen 3 14B on NVIDIA GeForce RTX 5080
- Sega Genesis Loads Games From a Vinyl Record in Viral Retro Hack
- Best Budget Local-AI Workstation Parts in 2026: 5 Picks
- IPEX-LLM + Ollama on Intel Arc: Setup, tok/s, and the RTX 3060 Reality Check
- Perplexity's Local-or-Cloud Router: What Hardware Runs the Local Half
- Ryzen 7 5800X vs Ryzen 7 5700X for Gaming and Local AI: Which Wins?
- Best Budget 4K Gaming Monitor for an RTX 3060: KOORUI vs Samsung Odyssey
- Noctua NH-U12S vs Cooler Master ML240L for Ryzen 5800X
More buying guides from SpecPicks
Browse all buying guides →- Best NVMe SSDs for Gaming in 2026
- Best External SSDs for Content Creators in 2026
- Best 4K Monitors for Content Creators in 2026
- Best NVMe External Enclosures for 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best Controllers for PC Gaming in 2026
- Best CPU Coolers for 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best CPUs for Gaming in 2026
- Best GPUs for 4K Gaming in 2026
- Best Graphics Cards for Gaming in 2026
- Best CPUs for Content Creators in 2026
- Best Retro Gaming Consoles & Handhelds for 2026
- Best Gaming Mice for 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best Gaming Monitors for 2026
- Best PC Cases for Building in 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best AM5 Motherboards for 2026
- Best GPUs for Running Local LLMs in 2026