96GB · AI Score 49.9/100 · first-party measured on 12 AI workloads
49.9 AI Score ✓ Measured
Every number on this page is first-party: NVIDIA RTX PRO 6000 Blackwell Server Edition was run on our pinned 12-workload AI suite on 2026-07-10, with under 0.5% run-to-run variance. On Llama 3.1 8B (Q4_K_M) NVIDIA RTX PRO 6000 Blackwell Server Edition delivers about 233.8 tokens/sec. Stepping up to Qwen3 32B it holds roughly 63.9 tok/s. The full Llama 3.3 70B still runs, at about 32.08 tok/s. For image generation, SDXL runs at 13.16 it/s, and FLUX.1-dev at 3.15 it/s. All 12 workloads fit in 96GB. There is no model in our suite this card has to turn down. NVIDIA RTX PRO 6000 Blackwell Server Edition isn't a retail purchase for most people. It's rented by the hour. You can run this exact card on RunPod.
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Qwen3 4B | 315.72 tok/s | 3 GB peak172 W43°C1.83 tok/WQ4_K_M | ✓ Measured |
| Llama 3.1 8B | 233.8 tok/s | 5.2 GB peak217 W45°C1.08 tok/WQ4_K_M | ✓ Measured |
| Qwen2.5-Coder 14B | 135.07 tok/s | 9 GB peak261 W48°C0.52 tok/WQ4_K_M | ✓ Measured |
| Qwen3 32B | 63.9 tok/s | 19 GB peak218 W51°C0.29 tok/WQ4_K_M | ✓ Measured |
| Llama 3.3 70B | 32.08 tok/s | 40.1 GB peak251 W55°C0.13 tok/WQ4_K_M | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Stable Diffusion XL | 26.32 images/min | 14.7 GB peak509 W59°C2.3 s/img | ✓ Measured |
| Z-Image Turbo | 14.78 images/min | 25.9 GB peak531 W63°C4.1 s/img | ✓ Measured |
| FLUX.1 dev | 6.75 images/min | 36.8 GB peak544 W67°C8.9 s/img | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| FLUX.1 Kontext dev | 3.02 images/min | 35.6 GB peak547 W72°C19.9 s/img | ✓ Measured |
| Qwen-Image-Edit | 2.62 images/min | 60.3 GB peak567 W73°C23 s/img | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| LTX-Video (distilled) | 15.64 frames/s | 60.4 GB peak524 W67°C6.2 s/clip | ✓ Measured |
| Wan 2.2 5B (720p) | 0.88 frames/s | 37.6 GB peak531 W72°C55.4 s/clip | ✓ Measured |
| Architecture | Blackwell |
| CUDA cores | 24,064 |
| VRAM | 96GB GDDR7 ECC |
| Memory bus | 512-bit |
| Memory bandwidth | 1792 GB/s |
| Boost clock | 2,610 MHz |
| TDP | 600 W |
| Process | TSMC 4N |
| Interface | PCIe 5.0 x16 |
| Release date | 2025-03-18 |
| Launch MSRP | $8,565 |
NVIDIA RTX PRO 6000 Blackwell Server Edition scores 49.9/100, #10 of 102. It ran all 12 workloads. Every figure here is our own measurement.
100% = this card, AI & Machine Learning headline metric (AI Score). #9 of 21 desktop cards in this vertical.
| GPU | Relative | % | AI Score |
|---|---|---|---|
| NVIDIA GH200 Grace Hopper | 131% | 65.6 | |
| NVIDIA H200 | 130% | 65 | |
| NVIDIA H100 80GB HBM3 | 125% | 62.6 | |
| NVIDIA H800 80GB | 125% | 62.6 | |
| NVIDIA RTX PRO 6000 Blackwell Server Edition | 100% | 49.9 | |
| NVIDIA H100 PCIe | 93% | 46.4 | |
| NVIDIA A100 80GB SXM4 | 66% | 33.1 | |
| NVIDIA A800 80GB | 66% | 33.1 | |
| NVIDIA A100 80GB PCIe | 64% | 31.7 |
← All AI & Machine Learning GPU rankings
Whole-job timings, composed from our measured per-model results on this card.
| Workflow | Time | Energy | Basis |
|---|---|---|---|
| 24-frame storyboard | 2.3 min | 15.7 Wh | all 2 stages measured |
| 60-second AI short film | 2.7 min | 19.22 Wh | all 3 stages measured |
| 6-panel comic page | 3.6 min | 26.83 Wh | all 3 stages measured |
| Character sheet, 12 poses | 4.6 min | 37.55 Wh | all 2 stages measured |
| Full codebase review | 7.4 min | 32.21 Wh | measured |
| 10 short social clips | 10.8 min | 88.63 Wh | all 3 stages measured |
| 20 long-form articles | 14.5 min | 60.83 Wh | measured |
| 40-product photo shoot | 15.2 min | 133.59 Wh | all 2 stages measured |
| 100-photo restoration batch | 33.4 min | 301.72 Wh | measured |