

NVIDIA H100 80GB HBM3 wins 10 of 17 benchmarks, averaging 7.7% faster.
NVIDIA H100 NVL's numbers are anchored estimates calibrated against our measured cards, pending first-party measurement. Treat small gaps as ties.
The gap is widest in Llama 3.3 70B, where NVIDIA H100 80GB HBM3 leads by 25% (41 vs 32.79 tok/s); the closest fight is Wan 2.2 5B (720p) (0% apart); VRAM decides part of this one: NVIDIA H100 NVL runs 201 of our 12 AI workloads while the other card runs 17, models that don't fit score zero.
| Benchmark | NVIDIA H100 80GB HBM3 | NVIDIA H100 NVL | Difference |
|---|---|---|---|
| Qwen3 4B tok/s | 310.26 | 282.91 | +10% |
| Llama 3.1 8B tok/s | 261.83 | 231.86 | +13% |
| Qwen3 8B tok/s | 244.22 | 221.81 | +10% |
| Gemma 4 12B tok/s | 148.17 | 135.74 | +9% |
| Qwen2.5-Coder 14B tok/s | 144.84 | 125.63 | +15% |
| Qwen3 14B tok/s | 151.16 | 135.54 | +12% |
| gpt-oss-20b tok/s | 346.8 | 296.75 | +17% |
| Qwen3 30B A3B tok/s | 283.78 | 270.25 | +5% |
| Qwen3 32B tok/s | 74.07 | 64.06 | +16% |
| Llama 3.3 70B tok/s | 41 | 32.79 | +25% |
| Stable Diffusion XL images/min | 34.58 | 34.58 | 0% |
| Z-Image Turbo images/min | 22.125 | 22.125 | 0% |
| FLUX.1 Kontext dev images/min | 4.264 | 4.264 | 0% |
| FLUX.1 dev images/min | 9.064 | 9.064 | 0% |
| Qwen-Image-Edit images/min | 3.68 | 3.68 | 0% |
| LTX-Video (distilled) frames/s | 17.47 | 17.47 | 0% |
| Wan 2.2 5B (720p) frames/s | 1.33 | 1.33 | 0% |
How long each card takes to finish a complete pipeline, not just one model. NVIDIA H100 80GB HBM3 is faster on 2 of 9; NVIDIA H100 NVL on 7.
| Workflow | NVIDIA H100 80GB HBM3 | NVIDIA H100 NVL | Difference | Cost per run |
|---|---|---|---|---|
| 24-frame storyboard | 1.8 min | 89 s | NVIDIA H100 NVL 1.22x faster | $0.081 vs $0.064 |
| 60-second AI short film | 2.5 min | 2.1 min | NVIDIA H100 NVL 1.19x faster | $0.111 vs $0.091 |
| 6-panel comic page | 3 min | 2.3 min | NVIDIA H100 NVL 1.30x faster | $0.133 vs $0.099 |
| Character sheet, 12 poses | 3.7 min | 2.9 min | NVIDIA H100 NVL 1.25x faster | $0.164 vs $0.126 |
| Full codebase review | 6.9 min | 8 min | NVIDIA H100 80GB HBM3 1.16x faster | $0.310 vs $0.345 |
| Short social clips | 8 min | 6.8 min | NVIDIA H100 NVL 1.18x faster | $0.358 vs $0.292 |
| Product photo shoot | 11.1 min | 10.5 min | NVIDIA H100 NVL 1.06x faster | $0.499 vs $0.455 |
| Long-form article batch | 11.4 min | 14.3 min | NVIDIA H100 80GB HBM3 1.26x faster | $0.510 vs $0.617 |
| Photo restoration batch | 23.8 min | 23.5 min | NVIDIA H100 NVL 1.02x faster | $1.068 vs $1.012 |
Renting by the hour, NVIDIA H100 NVL finishes 7 of 9 cheaper. The quicker card is not automatically the cheaper way to get the work done.
| Card | Per hour |
|---|---|
| NVIDIA H100 80GB HBM3 | $2.690 |
| NVIDIA H100 NVL | $2.590 |
NVIDIA H100 NVL is 1.04x cheaper per hour. Cheapest on-demand rate we see across RunPod and Vast.
| NVIDIA H100 80GB HBM3 | NVIDIA H100 NVL | |
|---|---|---|
| VRAM | 80GB | 94GB |
| Transistors | 80,000M | 80,000M |
| Die size | 814 mm² | 814 mm² |
| Process node | 4 nm | 4 nm |
| Transistor density | 98.3 M/mm² | 98.3 M/mm² |
| Architecture | Hopper | Hopper |
| Memory bandwidth | 3350 GB/s | 3938 GB/s |
| Boost clock | 1,980 MHz | 1,785 MHz |
| TDP | 700 W | 400 W |
| Launch MSRP | $30,000 | $29,000 |
| Release | 2022-09-20 | 2023-03-21 |
NVIDIA H100 80GB HBM3 full review · NVIDIA H100 NVL full review · All AI & Machine Learning rankings