

NVIDIA H200 wins 17 of 17 benchmarks, averaging 11.3% faster.
NVIDIA H100 NVL's numbers are anchored estimates calibrated against our measured cards, pending first-party measurement. Treat small gaps as ties.
The gap is widest in Llama-3.3-70B, where NVIDIA H200 leads by 30% (32.79 vs 42.66 tok/s); the closest fight is Qwen-Image-Edit (2% apart); VRAM decides part of this one: NVIDIA H100 NVL runs 164 of our 12 AI workloads while the other card runs 17, models that don't fit score zero.
| Benchmark | NVIDIA H100 NVL | NVIDIA H200 | Difference |
|---|---|---|---|
| Qwen3-4B tok/s | 282.91 | 318.84 | -11% |
| Llama-3.1-8B tok/s | 231.86 | 268.31 | -14% |
| Qwen3 8B tok/s | 221.81 | 247.98 | -11% |
| Gemma 4 12B tok/s | 135.74 | 151.05 | -10% |
| Qwen2.5-Coder-14B tok/s | 125.63 | 148.43 | -15% |
| Qwen3 14B tok/s | 135.54 | 154.51 | -12% |
| gpt-oss-20b tok/s | 296.75 | 354.23 | -16% |
| Qwen3 30B A3B tok/s | 270.25 | 292.11 | -7% |
| Qwen3-32B tok/s | 64.06 | 76.58 | -16% |
| Llama-3.3-70B tok/s | 32.79 | 42.66 | -23% |
| Stable Diffusion XL images/min | 34.58 | 37.16 | -7% |
| FLUX.1 Kontext dev images/min | 4.264 | 4.393 | -3% |
| FLUX.1 dev images/min | 9.064 | 9.514 | -5% |
| Qwen-Image-Edit images/min | 3.68 | 3.76 | -2% |
| Z-Image Turbo images/min | 22.125 | 23.175 | -5% |
| LTX-Video (distilled) frames/s | 17.47 | 17.87 | -2% |
| Wan 2.2 5B (720p) frames/s | 1.33 | 1.41 | -6% |
How long each card takes to finish a complete pipeline, not just one model. NVIDIA H100 NVL is faster on 6 of 9; NVIDIA H200 on 3.
| Workflow | NVIDIA H100 NVL | NVIDIA H200 | Difference | Cost per run |
|---|---|---|---|---|
| 24-frame storyboard | 89 s | 3.9 min | NVIDIA H100 NVL 2.62x faster | $0.064 vs $0.234 |
| 60-second AI short film | 2.1 min | 3.9 min | NVIDIA H100 NVL 1.86x faster | $0.091 vs $0.233 |
| 6-panel comic page | 2.3 min | 3.2 min | NVIDIA H100 NVL 1.40x faster | $0.099 vs $0.192 |
| Character sheet, 12 poses | 2.9 min | 3.9 min | NVIDIA H100 NVL 1.33x faster | $0.126 vs $0.234 |
| Short social clips | 6.8 min | 10 min | NVIDIA H100 NVL 1.48x faster | $0.292 vs $0.598 |
| Full codebase review | 8 min | 6.7 min | NVIDIA H200 1.19x faster | $0.345 vs $0.403 |
| Product photo shoot | 10.5 min | 11.2 min | NVIDIA H100 NVL 1.06x faster | $0.455 vs $0.668 |
| Long-form article batch | 14.3 min | 10.9 min | NVIDIA H200 1.31x faster | $0.617 vs $0.655 |
| Photo restoration batch | 23.5 min | 23.3 min | NVIDIA H200 1.00x faster | $1.012 vs $1.396 |
Renting by the hour, NVIDIA H100 NVL finishes 9 of 9 cheaper. The quicker card is not automatically the cheaper way to get the work done.
| Card | Per hour |
|---|---|
| NVIDIA H100 NVL | $2.590 |
| NVIDIA H200 | $3.590 |
NVIDIA H100 NVL is 1.39x cheaper per hour. Cheapest on-demand rate we see across RunPod and Vast.
| NVIDIA H100 NVL | NVIDIA H200 | |
|---|---|---|
| VRAM | 94GB | 141GB |
| Transistors | 80,000M | 80,000M |
| Die size | 814 mm² | 814 mm² |
| Process node | 4 nm | 4 nm |
| Transistor density | 98.3 M/mm² | 98.3 M/mm² |
| Architecture | Hopper | Hopper |
| Memory bandwidth | 3938 GB/s | 4800 GB/s |
| Boost clock | 1,785 MHz | 1,980 MHz |
| TDP | 400 W | 700 W |
| Launch MSRP | $29,000 | $31,000 |
| Release | 2023-03-21 | 2024-03-18 |
NVIDIA H100 NVL full review · NVIDIA H200 full review · All AI & Machine Learning rankings