NVIDIA Jetson Orin NX 8GB — Benchmarks & Specs
Bottom line: how fast is the NVIDIA Jetson Orin NX 8GB?
For local LLM inference it generates 7.5 tokens/sec running gemma4:12b at q4_K_M under llama.cpp, per NVIDIA Developer Forums. Its 8 GB of memory, shared by the CPU and GPU, is the binding constraint for local inference: the gemma4:12b run above (12B parameters) is the largest model on file at 4-bit quantization on it.
Every figure above is a row in the tables below, and each row links out to the review or public benchmark database the number was taken from; the memory size is the on-paper spec. SpecPicks aggregates published measurements; it does not report first-party benchmark runs.
The NVIDIA Jetson Orin NX 8GB is an APU from the Jetson Orin family released in 2023 from NVIDIA. Key on-paper specs include 6 CPU cores, 512 GPU cores, 8 GB of memory shared by the CPU and GPU. It launched with a $399 MSRP, though street prices typically diverge meaningfully from launch pricing. Data on this page draws on 16 community AI inference reports (top 12 shown), compiled from NVIDIA Developer Blog, Advantech-Containers, NVIDIA Developer Forums, ProventusNova; each table row links to its source. Read this page when shopping the NVIDIA Jetson Orin NX 8GB, comparing it against other APUs in your build, or sizing it for a specific workload (local LLM inference).
AI Inference Performance
Tokens per second under each model + quantization. Higher = faster generation. Bars compare runs across the same model.
| Model | Quantization | Relative | Tokens/sec | VRAM used | Source |
|---|---|---|---|---|---|
| smollm2:1.7b | INT4 mlc | 51.5 tok/s | — | NVIDIA Technical Blog 2025-01-16 | |
| llama3.2:3b | INT4 mlc | 34.5 tok/s | — | NVIDIA Technical Blog 2025-01-16 | |
| phi3.5-mini:3.8b | INT4 mlc | 30.8 tok/s | — | NVIDIA Technical Blog 2025-01-16 | |
| phi-3.5-mini:3.8b | int4 mlc-llm | 30.8 tok/s | — | NVIDIA Developer Blog 2025-01-01 | |
| gemma2:2b | INT4 mlc | 26.6 tok/s | — | NVIDIA Technical Blog 2025-01-16 | |
| llama3.2:1b | Q8_0 llama.cpp | 18.5 tok/s | 1.3 GB | Advantech-Containers (GitHub) 2025-05-01 | |
| qwen2.5:7b | INT4 mlc | 17.1 tok/s | — | NVIDIA Technical Blog 2025-01-16 | |
| llama3.2:1b | q4_K_M llama.cpp | 16.5 tok/s | 1.0 GB | Advantech-Containers (GitHub) 2025-05-01 | |
| qwen2.5:1.5b | q4_K_M llama.cpp | 16.0 tok/s | 1.0 GB | Advantech-Containers (GitHub) 2025-05-01 | |
| llama3.1:8b | INT4 mlc | 15.9 tok/s | — | NVIDIA Technical Blog 2025-01-16 | |
| gemma4:4b | q4_K_M llama.cpp | 13.0 tok/s | — | NVIDIA Developer Forums 2026-06-06 | |
| llama3.1:8b | q4_K_M llama.cpp | 11.0 tok/s | — | ProventusNova 2025-01-01 |
Full Specifications
| ram gb | 8 |
|---|---|
| ai tops | 70 |
| cpu cores | 6 |
| gpu cores | 512 |
NVIDIA Jetson Orin NX 8GB — Frequently Asked Questions
What is the NVIDIA Jetson Orin NX 8GB best used for?
When was the NVIDIA Jetson Orin NX 8GB released, and what was its launch MSRP?
Where do the NVIDIA Jetson Orin NX 8GB benchmark numbers come from?
How many cores does the NVIDIA Jetson Orin NX 8GB have, and how does it benchmark?
Where can I buy the NVIDIA Jetson Orin NX 8GB?
Buying guides that rank the NVIDIA Jetson Orin NX 8GB's class
This page is the raw performance data. The guides below turn it into a ranked pick for a specific build.
Editorial guides covering the NVIDIA Jetson Orin NX 8GB
In-depth SpecPicks reviews, build guides, and head-to-heads referencing this APU.
- 1080p Gaming PC Build 2026: Cost, Parts, FPS
- 4K Gaming PC Build 2026: Cost, Parts, FPS
- 1440p Gaming PC Build 2026: Cost, Parts, FPS
- Prime Big Deal Days 2026 CPU Deals: Ryzen 5600X to 9800X3D
- Prime Big Deal Days Prebuilt PCs: RTX 5060 vs 5060 Ti vs 5070
- Arc B580 12GB vs RTX 3060 12GB: Which 12GB Card for 1440p in 2026?
- Best Monitor Deals for Prime Big Deal Days 2026: 5 Picks for Gaming and AI Desks
- GTX 1660 VRAM: Why 8GB Doesn't Exist (Real Specs)
- Browse all SpecPicks reviews →
More guides & deep dives from the SpecPicks archive
Browse all articles & guides →- Best Retro Handhelds in 2026 — From $35 to $500
- Emulation Hardware in 2026: FPGA, Software, and Cart-Reader Ecosystems
- Best 1440p Gaming GPUs in 2026
- How to Build a Windows 98 Retro PC in 2026
- RTX 4070 Super vs RX 7800 XT — Which to Buy in 2026
- Best Budget Gaming PC Build 2026 — ~$1,000 ($800 on Sale)
- The Complete Voodoo5 5500 AGP Driver Guide (2026 Edition)
More reviews from the SpecPicks archive
Browse all reviews →- Noctua NH-U12S vs CoolerMaster ML240L: Air or AIO for 5800X
- Best Budget PC Gaming Accessories Under $100 in 2026
- A LEGO Castle Case for the Raspberry Pi 5 Is Going Viral — Here's the Build
- Ryzen 5 5600G vs Ryzen 7 5700X for a Budget Gaming PC
- Best Gaming CPU for 1080p and 1440p Builds in 2026
- Sound BlasterX G6 on a Windows 98/XP Retro Gaming PC: What It Actually Does
- 18 LLMs Benchmarked on OCR: Why Cheaper Models Often Win
- Best Budget GPU for Local LLMs Under $300 in 2026: Why the RTX 3060 12GB Still Wins
- UT99 OldUnreal 469 Patch in 2026: Migrate Configs, Rejoin Modern Servers, and Tune Mouse Precision
- Build a Raspberry Pi 4 NAS With an SSD in 2026
- Best Game Controller for Couch & Big-Screen PC Gaming in 2026
- Ryzen 7 5800X vs Core i7-9700K for a 24/7 Game Server: Which Hosts More Players?
- Best Mini PC for Local LLM Inference in 2026
- AI-Assisted Driver Hunting on Voodoo3 + GeForce 4 Ti: A 2026 Win98 Workflow
- AMD Ryzen 7 9800X3D Bundle: DDR5 + MSI B850 Wi-Fi 7 Guide
- Intel llm-scaler-vllm 1.4: Arc Pro B70 Inference Support Lands
- Best GPU for Qwen 3 14B (2026)
- Anthropic's Fable 5 Ban and Jailbreak: What It Means for Local-LLM Resilience
- A 4B Local Coding Agent That Hits 87%: Does It Run on a 3060?
- First Homelab Setup: Hardware, Security & Networking Guide
- Best SATA SSD for a Retro Windows 98 Build: BX500 vs 870 EVO
- Ethernet WiFi Router on a Pi Pico 2W: What's Possible
- Ryzen AI Max 400 Gorgon Halo vs RTX 3060 for Local LLMs
- Best Steam Deck Docked Setup for a 4K Desk in 2026: Dock, Monitor, Storage
More buying guides from SpecPicks
Browse all buying guides →- Best Gaming Monitors for 2026
- Best Mechanical Keyboards for Gaming in 2026
- Best CPUs for Gaming in 2026
- Best GPUs for 4K Gaming in 2026
- Best GPUs for Running Local LLMs in 2026
- Best GPU for Running 27B-32B Local LLMs in 2026
- Best NVMe SSDs for Gaming in 2026
- Best 4K Monitors for Content Creators in 2026
- Best CPU Coolers for 2026
- Best 1440p 240Hz Gaming Monitors in 2026
- Best NVMe External Enclosures for 2026
- Best Retro Gaming Consoles & Handhelds for 2026
- Best Gaming Mice for 2026
- Best CPUs for Content Creators in 2026
- Best DDR5 RAM for Gaming PCs in 2026
- Best Tools for Building and Repairing Retro PCs in 2026
- Best PC Cases for Building in 2026
- Best AM5 Motherboards for 2026
- Best Controllers for PC Gaming in 2026
- Best External SSDs for Content Creators in 2026
- Best Graphics Cards for Gaming in 2026
Hardware benchmark data on SpecPicks
All benchmarks →- RTX 4000 Ada Generation — benchmarks & specs
- GRID RTX6000-24Q — benchmarks & specs
- Radeon RX 6750 XT — benchmarks & specs
- Ryzen 5 5600X — benchmarks & specs
- Radeon RX 6700 XT — benchmarks & specs
- AMD Ryzen 9 7900 — benchmarks & specs
- Intel Arc B390 GPU — benchmarks & specs
- NVIDIA GeForce RTX 3060 — benchmarks & specs
- NVIDIA GeForce RTX 2080 Ti — benchmarks & specs
- Radeon RX 7500 — benchmarks & specs
- Radeon RX 6600 XT — benchmarks & specs
- AMD Instinct MI355X 288GB — benchmarks & specs
- AMD Ryzen 5 9500F — benchmarks & specs
- Ryzen 5 3600 — benchmarks & specs
- Ryzen 5 2600 — benchmarks & specs
- Radeon RX 7600 — benchmarks & specs
- AMD Ryzen 5 PRO 7640HS — benchmarks & specs
- Ryzen 7 2700 — benchmarks & specs
- AMD Ryzen Threadripper PRO 9975WX — benchmarks & specs
- Intel Core Ultra 5 245K — benchmarks & specs
- GRID RTX6000-8Q — benchmarks & specs
- Apple M4 Pro 14 Core — benchmarks & specs
- Intel Arc A310 LP — benchmarks & specs
- AMD Ryzen 5 5500H — benchmarks & specs