24GB · AI Score 7.4/100 · first-party measured on 12 AI workloads
7.4 AI Score Includes estimates
Every number on this page is first-party: NVIDIA Quadro RTX 6000 (Turing) was run on our pinned 12-workload AI suite on 2026-07-11, with under 0.5% run-to-run variance. On Llama 3.1 8B (Q4_K_M) NVIDIA Quadro RTX 6000 (Turing) delivers about 84.21 tokens/sec. Llama 3.3 70B does not fit. It needs roughly 42GB and this card has 24GB. For image generation, SDXL runs at 2.97 it/s, while FLUX.1-dev won't fit at BF16 (needs ~26GB). 5 of the 12 workloads won't fit on 24GB at the tested precision, Qwen3 32B, Llama 3.3 70B, FLUX.1-dev, FLUX.1 Kontext and others. We publish those as hard gates rather than quietly dropping to a smaller quant.
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Qwen3 4B | 125.23 tok/s | 2.9 GB peak153 W49°C0.82 tok/WQ4_K_M | ✓ Measured |
| Llama 3.1 8B | 84.21 tok/s | 4.8 GB peak174 W57°C0.49 tok/WQ4_K_M | ✓ Measured |
| Qwen2.5-Coder 14B | 46.54 tok/s | 8.5 GB peak174 W66°C0.27 tok/WQ4_K_M | ✓ Measured |
| Qwen3 32B | n/a tok/s | estimated | Est. |
| Llama 3.3 70B | ✕ Won't fit needs ~46 GB | VRAM-gated at this precision | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Stable Diffusion XL | 5.94 images/min | 12.7 GB peak241 W76°C10.1 s/img | ✓ Measured |
| FLUX.1 dev | ✕ Won't fit needs ~26 GB | VRAM-gated at this precision | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| FLUX.1 Kontext dev | ✕ Won't fit needs ~26 GB | VRAM-gated at this precision | ✓ Measured |
| Qwen-Image-Edit | ✕ Won't fit needs ~42 GB | VRAM-gated at this precision | ✓ Measured |
| Architecture | Turing (TU102) |
| CUDA cores | 4,608 |
| VRAM | 24GB GDDR6 |
| Memory bus | 384-bit |
| Memory bandwidth | 672 GB/s |
| Boost clock | 1,770 MHz |
| TDP | 260 W |
| Process | 12nm |
| Interface | PCIe 3.0 x16 |
| Release date | 2018-08-14 |
| Launch MSRP | $6,300 |
NVIDIA Quadro RTX 6000 (Turing) scores 7.4/100, #36 of 102. It ran 4 of 12; 5 exceeded its 24GB. Every figure here is our own measurement.
100% = this card, AI & Machine Learning headline metric (AI Score). #10 of 20 desktop cards in this vertical.
| GPU | Relative | % | AI Score |
|---|---|---|---|
| AMD Radeon Pro W7900 | 173% | 12.8 | |
| NVIDIA Quadro RTX 8000 | 111% | 8.2 | |
| NVIDIA RTX A5500 | 103% | 7.6 | |
| NVIDIA RTX PRO 4000 Blackwell | 103% | 7.6 | |
| NVIDIA Quadro RTX 6000 (Turing) | 100% | 7.4 | |
| NVIDIA RTX A5000 | 100% | 7.4 | |
| AMD Radeon Pro W7800 | 84% | 6.2 | |
| NVIDIA RTX 4500 Ada Generation | 80% | 5.9 | |
| AMD Radeon Pro W6800 | 74% | 5.5 |
← All AI & Machine Learning GPU rankings
Whole-job timings, composed from our measured per-model results on this card.
| Workflow | Time | Energy | Basis |
|---|---|---|---|
| Full codebase review | 21.5 min | 62.2 Wh | measured |
Can't run: 40-product photo shoot (needs FLUX.1 Kontext dev), 20 long-form articles (needs Llama 3.3 70B), Character sheet, 12 poses (needs FLUX.1 dev), 100-photo restoration batch (needs FLUX.1 Kontext dev), 100-photo restore and enlarge (needs FLUX.1 Kontext dev).