Skip to main content
RTX 4000 SFF Ada Generation
NVIDIA · GPU · Ada Lovelace

RTX 4000 SFF Ada Generation — Benchmarks & Specs

*Price sourced from Amazon.com. Price and availability subject to change.

Bottom line: how fast is the RTX 4000 SFF Ada Generation?

At 4K (High), the RTX 4000 SFF Ada Generation averages 144 fps in Apex Legends, per CpuTronic. For local LLM inference it generates 768.0 tokens/sec running qwen3:30b at Q4 under llama.cpp, per LocalLLaMA. In Geekbench 5 OpenCL it scores 124,926 points, per technical.city.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The RTX 4000 SFF Ada Generation is a graphics card from the Ada Lovelace family from NVIDIA. Data on this page draws on 1+ Amazon listings, 10 synthetic benchmark results, 12 community AI inference reports, 2 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the RTX 4000 SFF Ada Generation, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the RTX 4000 SFF Ada Generation by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Apex Legends 4K High 144 fps CpuTronic 2025-04-01
Cyberpunk 2077: Phantom Liberty 4K RT Ultra RT on DLSS Quality 68 fps CpuTronic 2025-04-01

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the RTX 4000 SFF Ada Generation, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
qwen3:30b Q4 llama.cpp 768.0 tok/s LocalLLaMA 2026-02-25
qwen3:235b Q4 ollama 324.0 tok/s LocalLLaMA 2026-02-27
llama3.2:1b q4_K_M ollama 189.0 tok/s LocalScore AI 2024-10-01
llama3.2:1b q4_K_M ollama 189.0 tok/s LocalScore (Mozilla Builders) 2025-04-14
llama3.2:1b q4_K_M llama.cpp 189.0 tok/s LocalScore 2024-06-01
llama3.2:1b q4_K_M llama.cpp 189.0 tok/s LocalScore.ai 2024-11-01
llama3.2:1b q4_K_M llama.cpp 189.0 tok/s localscore.ai 2024-01-01
llama3.1:8b q4_K_M llama.cpp 58.6 tok/s Hardware Corner 2024-09-29
llama3.1:8b q4_K_M llama.cpp 58.6 tok/s Hardware Corner 2024-09-29
llama3.1:8b q4_K_M llama.cpp 58.6 tok/s hardware-corner.net 2024-09-29
llama3.1:8b q4_K_M ollama 44.4 tok/s LocalScore (Mozilla Builders) 2025-04-14
llama3.1:8b q4_K_M llama.cpp 44.4 tok/s LocalScore.ai 2024-11-01

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the RTX 4000 SFF Ada Generation — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
Geekbench 5 OpenCL 124,926 points technical.city 2023-06-01
Geekbench 5 Vulkan 110,912 points technical.city 2023-06-01
PassMark G3D Mark 20,605 points PassMark VideoCardBenchmark 2023-04-21
PassMark G3D Mark 20,492 pts PassMark 2026-04-20
PassMark G3D Mark 20,492 points PassMark Software 2023-04-21
3DMark Time Spy 13,990 points 3DMark 2023-06-01
3DMark Steel Nomad (DX12 Graphics Score) 2,206 points UL Benchmarks 2024-11-01
3DMark Steel Nomad 2,206 points UL Benchmarks (3DMark) 2025-01-01
3DMark Steel Nomad 2,206 points UL Benchmarks (Futuremark) 2024-01-01
PassMark G2D Mark 1,081 pts PassMark 2026-04-20

Products Featuring the RTX 4000 SFF Ada Generation

RTX 4000 SFF Ada Generation — Frequently Asked Questions

What is the RTX 4000 SFF Ada Generation best used for?
RTX 4000 SFF Ada Generation is positioned as a Ada Lovelace-family graphics card. Use it for 1080p gaming and general productivity. See the synthetic + AI benchmark tables below for measured performance.
When was the RTX 4000 SFF Ada Generation released, and what was its launch MSRP?
Release year and launch MSRP aren't on file for RTX 4000 SFF Ada Generation. Current pricing is visible on the linked Amazon product cards below.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the RTX 4000 SFF Ada Generation run local LLMs?
Yes — RTX 4000 SFF Ada Generation has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers).
Where can I buy the RTX 4000 SFF Ada Generation?
Amazon listings are linked at the top of this page (in the hero CTA) and in the "Products Featuring this Hardware" section below. SpecPicks earns a small affiliate commission on qualifying purchases.

Buying guides that rank the RTX 4000 SFF Ada Generation's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the RTX 4000 SFF Ada Generation

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →
RTX 4000 SFF Ada Generation
RTX 4000 SFF Ada Generation
View on Amazon →