Apple M2 Ultra 24 Core — Benchmarks & Specs
Bottom line: how fast is the Apple M2 Ultra 24 Core?
For local LLM inference it generates 94.3 tokens/sec running llama2:7b at q4_0 under llama.cpp, per llama.cpp GitHub. In PassMark CPU Mark it scores 50,950 points, per PassMark.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The Apple M2 Ultra 24 Core is a processor from Apple. Data on this page draws on 10 synthetic benchmark results, 12 community AI inference reports, aggregated from public benchmark databases (TechPowerUp, PassMark, Geekbench, Cinebench) and the LocalLLaMA community. Read this page when shopping the Apple M2 Ultra 24 Core, comparing it against other processors in your build, or sizing it for a specific workload (gaming at 1080p/1440p/4K, productivity benchmarks, or local LLM inference).
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| llama2:7b | q4_0 llama.cpp | 94.3 tok/s | — | llama.cpp GitHub 2023-08-01 | |
| llama:7b | q4_0 llama.cpp | 94.3 tok/s | — | llama.cpp GitHub Discussion #4167 2023-11-22 | |
| llama2:7b | q4_0 llama.cpp | 88.6 tok/s | — | llama.cpp GitHub Discussions 2023-10-01 | |
| llama3:8b | q4_K_M llama.cpp | 76.3 tok/s | — | XiongjieDai GPU-Benchmarks-on-LLM-Inference (GitHub) 2024-05-01 | |
| llama3:8b | q4_K_M llama.cpp | 76.3 tok/s | — | GitHub/XiongjieDai GPU-Benchmarks-on-LLM-Inference 2024-05-01 | |
| llama2:7b | q8_0 llama.cpp | 66.6 tok/s | — | llama.cpp GitHub 2023-08-01 | |
| llama:7b | q8_0 llama.cpp | 66.6 tok/s | — | llama.cpp GitHub Discussion #4167 2023-11-22 | |
| llama2:7b | q8_0 llama.cpp | 62.1 tok/s | — | llama.cpp GitHub Discussions 2023-10-01 | |
| llama2:7b | FP16 llama.cpp | 41.0 tok/s | — | llama.cpp GitHub 2023-08-01 | |
| llama:7b | FP16 llama.cpp | 41.0 tok/s | — | llama.cpp GitHub Discussion #4167 2023-11-22 | |
| mixtral:8x22b | q4_K_M mlx | 18.0 tok/s | — | Kunal Ganglani LLM Benchmarks 2024-08-01 | |
| llama3.1:70b | q5_K_M mlx | 13.0 tok/s | — | Kunal Ganglani LLM Benchmarks 2024-08-01 |
Synthetic Benchmarks
Higher is better. Bars are scaled within each benchmark family (multi-thread, single-thread, etc.) so you can compare like-with-like at a glance.
| Benchmark | Relative | Score | Source |
|---|---|---|---|
| PassMark CPU Mark | 50,950 points | PassMark 2023-09-01 | |
| PassMark CPU Mark | 50,892 points | PassMark Software 2026-06-02 | |
| Cinebench R23 multi | 28,730 points | XDA-Developers 2023-06-17 | |
| Cinebench R23 multi | 28,665 points | Engadget 2023-06-12 | |
| Cinebench R23 multi | 27,095 points | Tom's Hardware 2023-06-05 | |
| Geekbench 6 multi | 21,453 points | AppleInsider 2023-06-13 | |
| Geekbench 6 multi | 21,413 points | Geekbench Browser 2023-06-01 | |
| PassMark Single Thread | 4,208 points | PassMark Software 2026-09-04 | |
| Geekbench 6 single | 2,794 points | AppleInsider 2023-06-13 | |
| Geekbench 6 single | 2,776 points | Geekbench Browser 2023-06-01 |
Apple M2 Ultra 24 Core — Frequently Asked Questions
What is the Apple M2 Ultra 24 Core best used for?
When was the Apple M2 Ultra 24 Core released, and what was its launch MSRP?
Where do the benchmark numbers on this page come from?
How does the Apple M2 Ultra 24 Core compare to its predecessor?
Where can I buy the Apple M2 Ultra 24 Core?
Buying guides that rank the Apple M2 Ultra 24 Core's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the Apple M2 Ultra 24 Core
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this processor.
- Arc B580 12GB vs RTX 3060 12GB: Which 12GB Card for 1440p in 2026?
- One Monitor, Two PCs: Monitor KVM vs Dual-Input for a Gaming Rig and AI Box
- The Greatest GPU Awards: A Decade of Standout Cards
- Best Wired Gear for Tournament-Legal LAN Play in 2026
- GTX 1660 VRAM: Why 8GB Doesn't Exist (Real Specs)
- Alien: Isolation on Steam Deck: Best Settings for 2025
- Noctua NH-U12S vs Corsair H150i vs Kraken M22 on a 24/7 Ryzen 7 5800X Host
- How to Filter Steam Deck's 20K+ Games by FPS, Price, HLTB
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Best Retro Handhelds in 2026 — From $35 to $500
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- How to Build a Windows 98 Retro PC in 2026
- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- Best 1440p Gaming GPUs in 2026
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
More reviews from the SpecPicks archive
Browse all reviews →- Best 2.5" SATA SSD for Old Laptop and PC Upgrades in 2026
- Ryzen 7 5800X vs Core i7-9700K for a 24/7 Game Server: Which Hosts More Players?
- RTX 3060 12GB vs Ryzen 5600G iGPU: Best 1080p Entry Path
- Intel Arc Pro B60 Gaming Performance: 2025 Benchmarks
- Intel Axes BigDL/IPEX-LLM: Where Local Inference Goes Now
- Best SSD for Steam Deck OLED Expansion in 2026: SATA Adapter vs M.2 2230
- Build a Pi Zero 2 W Retro Mini-TV: Full Parts List and Setup
- Raspberry Pi 4 8GB Cyberdeck Build: A Practical 2026 Parts Guide
- Dual RTX 3090 vs RTX 5090: Gaming vs AI Training
- Cosmos3-Super on an RTX 3060 12GB: Can the #1 Open-Weights Image Model Run Local?
- Best SSD for a 4TB Steam Game Library Under $250 (2026)
- vLLM vs llama.cpp for Single-User Chat on an RTX 3060 12GB (2026)
- CompactFlash as Your IDE Hard Drive: A Silent, Reliable Retro PC Boot Disk
- retrorecs.com: A New Retro Game Recommendation Engine
- Run Qwen Locally: Apple Silicon vs a 12GB RTX 3060 Rig in 2026
- Best Game Controller for PC in 2026: 5 Picks From DualSense to Arcade Stick
- KOORUI 27" 4K vs Samsung Odyssey 4K: Best 27-Inch Panel for PC and Console
- Jetson Orin Nano Super for Local LLM: 7B/13B Tokens-per-Second Reality Check
- Best Budget SSDs for Retro and Homelab Builds in 2026
- Hyperscalers May Soon Outrun Their Cash Flow on AI Buildout
- The Hackaday Europe 2026 Retro PC Build: What They Actually Built and Why It Matters
- Claude Opus 4.8 Tops the Intelligence Index: Cloud vs Local on a 3060
- Best Streaming Setup Gear for Beginners in 2026
- Using an LLM to Fix Win98 Voodoo & TNT Driver Installs
More buying guides from SpecPicks
Browse all buying guides →- Best Graphics Cards for Gaming in 2026
- Best CPUs for Gaming in 2026
- Best GPUs for Running Local LLMs in 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best 4K Monitors for Content Creators in 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best PC Cases for Building in 2026
- Best Gaming Monitors for 2026
- Best Gaming Mice for 2026
- Best NVMe SSDs for Gaming in 2026
- Best CPU Coolers for 2026
- Best Retro Gaming Consoles & Handhelds for 2026
- Best External SSDs for Content Creators in 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best CPUs for Content Creators in 2026
- Best GPUs for 4K Gaming in 2026
- Best NVMe External Enclosures for 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best Controllers for PC Gaming in 2026
- Best AM5 Motherboards for 2026