NVIDIA · Hopper
NVIDIA H200 SXM 141GB
141 GB VRAM700W TDP2024
Full NVIDIA H200 SXM 141GB benchmarks →
Side-by-side comparison of 2 GPU parts: NVIDIA H200 SXM 141GB vs NVIDIA B200. All released in 2024. Specs, synthetic benchmarks, AI inference tok/s, and Amazon availability — click into any column header for the per-SKU benchmark page.
*Price sourced from Amazon.com. Price and availability subject to change.
| Spec | NVIDIA H200 SXM 141GB | NVIDIA B200 |
|---|---|---|
| vram gb | 141 | 192 |
| vram type | HBM3e | HBM3e |
| tdp w | 700 | 1000 |
| cuda cores | 16896 | 16896 |
| tensor cores | 528 | — |
| l2 cache mb | — | 50 |
| process nm | 4 | 5 |
Higher is better. Per-row scores come from public benchmark databases (PassMark, TechPowerUp, Geekbench, Cinebench). Empty cells mean we don't yet have that benchmark on file for the SKU.
| Benchmark | NVIDIA H200 SXM 141GB | NVIDIA B200 |
|---|---|---|
| 3DMark Time Spy | 15,200 points | — |
| MLPerf Inference v4.0 — Llama 2 70B Offline Scenario | 31,712 tokens/s | — |
| MLPerf Inference v5.0 — Llama 2 70B Server Scenario | 33,000 tokens/s | — |
| MLPerf Inference v5.1 — Llama 2 70B Offline | — | 102,725 tok/s |
| MLPerf Inference v6.0 — DeepSeek R1 Offline | — | 58,582 tok/s |
| MLPerf Inference v6.0 — gpt-oss 120B Offline | — | 85,921 tok/s |
| Port Royal (RT) | 11,800 points | — |
Top community-reported tokens-per-second results for popular LLMs. See the per-SKU benchmark page for the full set including quantization, prompt length, and source citations.