Skip to main content
Quadro RTX 4000 (Mobile)
NVIDIA · GPU · Ada Lovelace

Quadro RTX 4000 (Mobile) — Benchmarks & Specs

Bottom line: how fast is the Quadro RTX 4000 (Mobile)?

At 1440p (Ultra High), the Quadro RTX 4000 (Mobile) averages 102 fps in Assassin's Creed Valhalla, per Digital Trends. For local LLM inference it generates 61.8 tokens/sec running llama2:7b at q4_0 under llama.cpp, per llama.cpp GitHub Discussion #15013. In 3DMark Fire Strike Graphics it scores 40,179 points, per Notebookcheck.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The Quadro RTX 4000 (Mobile) is a graphics card from the Ada Lovelace family from NVIDIA. Data on this page draws on 8 synthetic benchmark results, 5 community AI inference reports, 2 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the Quadro RTX 4000 (Mobile), comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the Quadro RTX 4000 (Mobile) by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Assassin's Creed Valhalla 1440p Ultra High 102 fps Digital Trends 2025-01-01
Cyberpunk 2077 1440p Ultra FSR Quality 74 fps Digital Trends 2025-01-01

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the Quadro RTX 4000 (Mobile), in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
llama2:7b q4_0 llama.cpp 61.8 tok/s llama.cpp GitHub Discussion #15013 2025-08-01
llama3.1:8b q4_K_M llama.cpp 58.6 tok/s XiongjieDai GPU-Benchmarks-on-LLM-Inference GitHub 2024-05-01
llama3.1:8b q4_K_M ollama 53.5 tok/s LocalScore.ai 2024-06-01
llama3.1:8b FP16 llama.cpp 20.9 tok/s XiongjieDai GPU-Benchmarks-on-LLM-Inference GitHub 2024-05-01
qwen2.5:32b IQ4_XS llama.cpp 16.4 tok/s Hardware Corner 2024-01-01

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the Quadro RTX 4000 (Mobile) — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
3DMark Fire Strike Graphics 40,179 points Notebookcheck 2024-10-25
3DMark Fire Strike 32,984 points Notebookcheck 2024-01-01
3DMark Time Spy 16,013 points Notebookcheck 2024-01-01
PassMark G3D Mark 11,692 pts PassMark 2026-04-20
PassMark G3D Mark 11,692 points PassMark Software 2026-05-09
3DMark Speed Way 4,034 points Notebookcheck 2024-01-01
3DMark Steel Nomad 3,624 points Notebookcheck 2024-01-01
PassMark G2D Mark 506 pts PassMark 2026-04-20

Quadro RTX 4000 (Mobile) — Frequently Asked Questions

What is the Quadro RTX 4000 (Mobile) best used for?
Quadro RTX 4000 (Mobile) is positioned as a Ada Lovelace-family graphics card. Use it for 1080p gaming and general productivity. See the synthetic + AI benchmark tables below for measured performance.
When was the Quadro RTX 4000 (Mobile) released, and what was its launch MSRP?
Release year and launch MSRP aren't on file for Quadro RTX 4000 (Mobile). Current pricing is visible on the linked Amazon product cards below.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the Quadro RTX 4000 (Mobile) run local LLMs?
Yes — Quadro RTX 4000 (Mobile) has 5 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers).
Where can I buy the Quadro RTX 4000 (Mobile)?
Active Amazon listings aren't on file for this exact SKU yet. See the linked benchmark sources and the Compare tool for adjacent parts that may be in stock — and check the /benchmarks index for the latest curated picks in this category.

Buying guides that rank the Quadro RTX 4000 (Mobile)'s class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the Quadro RTX 4000 (Mobile)

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →