Skip to main content
NVIDIA RTX A5000 24GB
NVIDIA · GPU · Ampere Pro

NVIDIA RTX A5000 24GB — Benchmarks & Specs

24 GB VRAM230W TDP$1,999 MSRP2021

Bottom line: how fast is the NVIDIA RTX A5000 24GB?

At 1440p (Highest), the NVIDIA RTX A5000 24GB averages 138 fps in Shadow of the Tomb Raider, per CpuTronic. For local LLM inference it generates 3274.3 tokens/sec running deepseek-r1-distill-qwen:1.5b at FP16 under vllm, per DatabaseMart — A5000 vLLM Benchmark. In Geekbench 5 CUDA it scores 190,987 points, per Technical.city. Its 24 GB of VRAM is the binding constraint for local inference: that capacity fits 32B-parameter models at Q4 without offloading to system RAM.

Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.

The NVIDIA RTX A5000 24GB is a graphics card from the Ampere Pro family released in 2021 from NVIDIA. Key on-paper specs include 24 GB of GDDR VRAM, 230W TDP. It launched with a $1,999 MSRP, though street prices typically diverge meaningfully from launch pricing — see the linked product cards below for current Amazon listings. Data on this page draws on 10 synthetic benchmark results, 12 community AI inference reports, 3 measured game frame-rate results, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the NVIDIA RTX A5000 24GB, comparing it against other graphics cards in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).

Gaming Performance (measured FPS)

Average and 1% low frame rates by game, resolution, and quality preset. Bars are scaled against the fastest result on this page.

Measured gaming frame rates for the NVIDIA RTX A5000 24GB by game, resolution, and quality preset. “1% low” is the frame-time floor that determines perceived smoothness. Each row links to its original review or benchmark database.
Game Resolution Settings Relative Avg FPS 1% low Source
Shadow of the Tomb Raider 1080p Highest 194 fps CpuTronic 2024-01-01
Shadow of the Tomb Raider 1440p Highest 138 fps CpuTronic 2024-01-01
Shadow of the Tomb Raider 4K Highest 73 fps CpuTronic 2024-01-01

AI Inference Performance

Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.

Local LLM inference throughput on the NVIDIA RTX A5000 24GB, in generated tokens per second. Higher is better; each row links to the community report or benchmark database it came from.
Model Quantization Relative Tokens/sec VRAM used Source
deepseek-r1-distill-qwen:1.5b FP16 vllm 3274.3 tok/s 3.4 GB DatabaseMart — A5000 vLLM Benchmark 2025-01-01
qwen2.5:3b FP16 vllm 2303.5 tok/s 5.8 GB DatabaseMart — A5000 vLLM Benchmark 2025-01-01
deepseek-r1-distill-qwen:7b FP16 vllm 1349.2 tok/s 15.0 GB DatabaseMart — A5000 vLLM Benchmark 2025-01-01
deepseek-r1-distill-llama:8b FP16 vllm 1285.6 tok/s 15.0 GB DatabaseMart — A5000 vLLM Benchmark 2025-01-01
qwen2.5-vl:7b FP16 vllm 1078.3 tok/s 16.0 GB DatabaseMart — A5000 vLLM Benchmark 2025-01-01
gemma-2-9b-it FP16 vllm 166.0 tok/s 18.0 GB DatabaseMart — A5000 vLLM Benchmark 2025-01-01
llama2:7b q4_0 llama.cpp 138.7 tok/s knightli.com (llama.cpp CUDA scoreboard) 2026-04-23
llama2:7b Q4_0 llama.cpp 138.7 tok/s KnightLi llama.cpp GPU Benchmark Scoreboard 2026-04-23
llama2:7b q4_0 llama.cpp 135.8 tok/s llama.cpp GitHub (CUDA scoreboard) 2026-04-23
llama2:7b Q4_0 llama.cpp 130.1 tok/s llama.cpp GitHub Discussion #15013 2024-01-01
llama4:scout Q4_K_M ollama 119.5 tok/s hkadm/ollama_gpu_test (GitHub) 2025-04-01
llama2:13b Q4_0 ollama 60.5 tok/s 14.4 GB DatabaseMart — Ollama A5000 GPU Benchmark 2025-01-01

Synthetic Benchmarks

Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.

Synthetic benchmark scores for the NVIDIA RTX A5000 24GB — higher is better. Each row links to the public database the number was taken from.
Benchmark Relative Score Source
Geekbench 5 CUDA 190,987 points Technical.city 2024-01-01
Geekbench 5 OpenCL 155,952 points Technical.city 2024-01-01
Geekbench 5 Vulkan 137,639 points Technical.city 2024-01-01
3DMark Fire Strike 27,271 points GPU Monkey 2022-01-01
PassMark G3D Mark 23,034 points PassMark VideoCardBenchmark 2025-01-01
PassMark G3D Mark 22,887 points PassMark 2021-05-07
3DMark Time Spy 14,600 points TopCPU GPU Ranking — 3DMark Time Spy 2024-01-01
3DMark Time Spy 14,182 points CpuTronic GPU Database 2024-01-01
3DMark Time Spy 14,182 points CpuTronic 2022-01-01
3DMark Steel Nomad 3,790 points UL Benchmarks (3DMark) 2026-07-09

Full Specifications

tdp w230
vram gb24
cuda cores8192
memory typeGDDR6

NVIDIA RTX A5000 24GB — Frequently Asked Questions

What is the NVIDIA RTX A5000 24GB best used for?
NVIDIA RTX A5000 24GB is positioned as a 24 GB VRAM Ampere Pro-family graphics card. Use it for high-end 4K gaming and local LLM inference. See the synthetic + AI benchmark tables below for measured performance.
When was the NVIDIA RTX A5000 24GB released, and what was its launch MSRP?
NVIDIA RTX A5000 24GB launched in 2021 at a $1,999 MSRP. Street prices diverge from launch pricing over a product's lifetime — check the linked Amazon listings on this page for current availability.
Where do the benchmark numbers on this page come from?
Synthetic benchmarks are scraped from public databases (TechPowerUp, PassMark, Geekbench Browser, Cinebench leaderboards). AI inference numbers come from the LocalLLaMA community (Reddit threads, llama.cpp / Ollama discussion logs, and Phoronix when available). Every benchmark row carries an inline source citation — click through to verify the original number.
Can the NVIDIA RTX A5000 24GB run local LLMs?
Yes — NVIDIA RTX A5000 24GB has 12 AI inference benchmarks on file (see the AI Inference Performance section above for model + tokens-per-second numbers). With 24 GB VRAM, it fits the popular 32B-parameter open-weight models at Q4 quantization comfortably.
Where can I buy the NVIDIA RTX A5000 24GB?
Active Amazon listings aren't on file for this exact SKU yet. See the linked benchmark sources and the Compare tool for adjacent parts that may be in stock — and check the /benchmarks index for the latest curated picks in this category.

Buying guides that rank the NVIDIA RTX A5000 24GB's class

This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.

Editorial guides covering the NVIDIA RTX A5000 24GB

In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.

More guides & deep dives from the SpecPicks archive

Browse all articles & guides →

More reviews from the SpecPicks archive

Browse all reviews →

More buying guides from SpecPicks

Browse all buying guides →

Hardware benchmark data on SpecPicks

All benchmarks →