NVIDIA Tesla P40 24GB — Benchmarks & Specs
Bottom line: how fast is the NVIDIA Tesla P40 24GB?
For local LLM inference it generates 29.2 tokens/sec running qwen3-30b-a3b (a mixture-of-experts model with ~3B active parameters, so dense models of similar size run far slower) at q4_K_M under llama.cpp, per LocalScore.ai. In Geekbench OpenCL it scores 62,287 points, per Geekbench Browser / askgeek.io. Its 24 GB of VRAM is the binding constraint for local inference: the qwen3-30b-a3b run above (30.5B parameters) is the largest model on file at 4-bit quantization on this card.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The NVIDIA Tesla P40 24GB is a graphics card from the Pascal Pro family released in 2016 from NVIDIA. Key on-paper specs include 24 GB of GDDR5 VRAM, 250W TDP. It launched with a $6,800 MSRP, though street prices typically diverge meaningfully from launch pricing. Data on this page draws on 10 synthetic benchmark results, 21 community AI inference reports (top 12 shown), compiled from LocalScore.ai, TinyComputers.io, TopCPU, 3DMark, Geekbench Browser, Geekbench Browser / askgeek.io, KnightLi llama.cpp GPU Benchmark Scoreboard, Like2Byte Tesla P40 LLM Guide, llama.cpp GitHub, PassMark Software; each table row links to its source. Read this page when shopping the NVIDIA Tesla P40 24GB, comparing it against other graphics cards in your build, or sizing it for a specific workload (synthetic benchmark scores or local LLM inference).
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| llama3.2:3b | — ollama | 94.3 tok/s | — | TinyComputers.io 2026-03-01 | |
| llama3.2:1b | q4_K_M llama.cpp | 90.5 tok/s | — | LocalScore.ai 2025-01-01 | |
| llama2:7b | q4_0 llama.cpp | 54.7 tok/s | — | llama.cpp GitHub Discussions 2024-12-01 | |
| llama-2:7b | Q4_0 llama.cpp | 54.7 tok/s | — | llama.cpp GitHub 2025-08-01 | |
| llama2:7b | q4_0 llama.cpp | 54.7 tok/s | — | KnightLi llama.cpp GPU Benchmark Scoreboard 2026-04-23 | |
| qwen2.5:7b | — ollama | 52.7 tok/s | — | TinyComputers.io 2026-03-01 | |
| llama3.1:8b | — ollama | 47.8 tok/s | — | TinyComputers.io 2026-03-01 | |
| mistral:7b | q4_K_M llama.cpp | 45.0 tok/s | — | Like2Byte Tesla P40 LLM Guide 2026-02-20 | |
| llama2:7b | q4_0 llama.cpp | 40.9 tok/s | — | LocalScore.ai 2024-06-01 | |
| llama-2:7b | Q4_0 llama.cpp | 40.9 tok/s | — | LocalScore.ai 2025-01-01 | |
| qwen3:30b-a3b | q4_K_M llama.cpp | 29.2 tok/s | — | LocalScore.ai 2025-01-01 | |
| qwen3-30b-a3b | q4_K_M llama.cpp | 29.2 tok/s | — | LocalScore.ai 2025-06-01 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| Geekbench OpenCL | 62,287 points | Geekbench Browser / askgeek.io 2024-01-01 | |
| Geekbench 5 OpenCL | 62,287 points | Geekbench Browser 2021-07-16 | |
| PassMark G3D Mark | 11,596 points | PassMark Software 2026-06-09 | |
| 3DMark Time Spy | 8,079 points | 3DMark 2023-01-01 | |
| 3DMark Time Spy | 8,079 points | TopCPU 2024-01-01 | |
| 3DMark Time Spy | 7,418 points | 3DMark (UL Benchmarks) 2020-01-01 | |
| PassMark GPU Compute | 4,082 Ops/Sec | PassMark PerformanceTest 2026-04-30 | |
| 3DMark Time Spy Extreme | 3,894 points | TopCPU 2024-01-01 | |
| Blender | 797 points | Blender Benchmark / topcpu.net 2023-01-01 | |
| OctaneBench | 166 points | OctaneBench / topcpu.net 2023-01-01 |
Full Specifications
| tdp w | 250 |
|---|---|
| vram gb | 24 |
| cuda cores | 3840 |
| memory type | GDDR5 |
NVIDIA Tesla P40 24GB — Frequently Asked Questions
What is the NVIDIA Tesla P40 24GB best used for?
When was the NVIDIA Tesla P40 24GB released, and what was its launch MSRP?
Where do the NVIDIA Tesla P40 24GB benchmark numbers come from?
Can the NVIDIA Tesla P40 24GB run local LLMs?
Where can I buy the NVIDIA Tesla P40 24GB?
Buying guides that rank the NVIDIA Tesla P40 24GB's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the NVIDIA Tesla P40 24GB
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this graphics card.
- Prime Big Deal Days 2026: Best Mini PC Deals for Local AI and Home Labs
- Raspberry Pi 4 8GB vs Ryzen 5 2600 for Gemma 3 4B: Which Cheap Box Wins?
- RTX 4090 vs RTX 5090 for Local LLM Inference: 24 GB vs 32 GB (2026)
- Ryzen 5 2600 vs Ryzen 5 5600X for CPU-Only gpt-oss 20B (2026)
- Ryzen 5 5600X vs Ryzen 5 5600G for CPU-Only Gemma 3 12B
- Best Hardware for Local OCR and Document AI in 2026
- Prime Big Deal Days 2026 GPU Deals for Local LLMs: 12GB vs 16GB Cards
- Best GPU for gpt-oss 20B in 2026
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Best 1440p Gaming GPUs in 2026
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- How to Build a Windows 98 Retro PC in 2026
- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
- Best Retro Handhelds in 2026 — From $35 to $500
More reviews from the SpecPicks archive
Browse all reviews →- GPU-Accelerated Autorouter Handles Monstrous PCB Designs
- Best GPU for 1440p Local Image Generation in 2026: Why the RTX 3060 12GB Still Wins on Value
- Best CPU Cooler for AM4 + Ryzen 5000 in 2026
- GPT-5.6 Sol vs Local Open-Weights: Why a 12GB Rig Still Earns Its Keep
- Building a Period-Correct 2001 GeForce 3 + Windows 98 SE Gaming Rig
- Intel Core Ultra 5 250K Plus Takes On the Ryzen 5 7600X3D for Mid-Range Gaming
- Qwen-Image-3.0 on an RTX 3060 12GB: Local Text-in-Image Gen
- CRT PC Monitors for Retro Gaming: 2026 Buying Guide
- Genesis Mini vs SNES Classic vs Raspberry Pi 4: Which Retro Setup Wins in 2026?
- Ryzen 7 9800X3D vs Ryzen 9 9950X3D: The 2026 X3D Buyer's Verdict
- SATA vs NVMe for a Ryzen 5800X Gaming Build: Does It Matter?
- Asus ROG Harpe II Extreme: A 24K-Gold, 65K-Sensor Gaming Mouse
- Build a RetroPie Handheld in 2026: A Complete Step-by-Step Guide
- Gemma 4 Tool-Calling Fix: Re-test Function Calls Locally
- Nous Hermes Desktop: A Local AI Agent for Your Own Hardware
- Best Budget Streaming & Podcast Gear in 2026
- ExLlamaV2 vs llama.cpp for Single-User Chat on an RTX 3060 12GB in 2026
- ZOTAC RTX 3060 12GB vs Intel Arc B580 12GB for 1440p Gaming in 2026
- Silicon Motion PCIe 6.0: Nvidia AI Drives Consumer Storage
- Best Single-Board Computer Project Kits for Beginners (2026)
- Best GPU for Llama 70B Local Inference in 2026: RTX 3060 12GB Dual vs RTX 3090 vs Gorgon Halo
- Best Mouse and Mousepad for FPS Aim Training in 2026
- GLM-5.2 for Local Agents: Can a 12GB RTX 3060 Run Long-Horizon Tasks?
- Best Game Controller in 2026: 5 Picks for PC, Console and Retro
More buying guides from SpecPicks
Browse all buying guides →- Best Gaming Monitors for 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best CPU Coolers for 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best Retro Gaming Consoles & Handhelds for 2026
- Best Controllers for PC Gaming in 2026
- Best CPUs for Gaming in 2026
- Best Gaming Mice for 2026
- Best NVMe SSDs for Gaming in 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best 4K Monitors for Content Creators in 2026
- Best GPUs for Running Local LLMs in 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best NVMe External Enclosures for 2026
- Best PC Cases for Building in 2026
- Best External SSDs for Content Creators in 2026
- Best CPUs for Content Creators in 2026
- Best Graphics Cards for Gaming in 2026
- Best AM5 Motherboards for 2026
- Best GPUs for 4K Gaming in 2026
- Best Tools for Building and Repairing Retro PCs in 2026
Hardware benchmark data on SpecPicks
All benchmarks →- Intel Arc A770 — benchmarks & specs
- AMD Ryzen 5 5600H — benchmarks & specs
- Radeon RX 6550S — benchmarks & specs
- NVIDIA GeForce RTX 3070 Ti — benchmarks & specs
- GeForce RTX 4080 — benchmarks & specs
- AMD Ryzen 9 7945HX3D — benchmarks & specs
- NVIDIA RTX PRO 6000 Blackwell — benchmarks & specs
- Apple M4 Max — benchmarks & specs
- AMD Ryzen 3 PRO 5355GE — benchmarks & specs
- Ryzen 7 9700X — benchmarks & specs
- Apple M2 Max 12 Core 3680 MHz — benchmarks & specs
- GRID RTX6000-1B — benchmarks & specs
- AMD Ryzen 3 7320C — benchmarks & specs
- AMD Ryzen Threadripper PRO 5955WX — benchmarks & specs
- Ryzen 7 5800H — benchmarks & specs
- GRID RTX6000P-6Q — benchmarks & specs
- AMD Ryzen Threadripper PRO 9945WX — benchmarks & specs
- Pentium 4 2.53GHz (Northwood) — benchmarks & specs
- DDR5-6000 CL30 64GB (2x32) — benchmarks & specs
- Radeon RX 6600 LE — benchmarks & specs
- Radeon 9500 Pro — benchmarks & specs
- Ryzen 3 2200G — benchmarks & specs
- Intel Arc B370 GPU — benchmarks & specs
- GeForce RTX 4060 Ti — benchmarks & specs