48GB · AI Score 23.8/100 · first-party measured on 12 AI workloads
23.8 AI Score ✓ Measured
Every number on this page is first-party: NVIDIA RTX 6000 Ada Generation was run on our pinned 12-workload AI suite on 2026-07-12, with under 0.5% run-to-run variance. On Llama 3.1 8B (Q4_K_M) NVIDIA RTX 6000 Ada Generation delivers about 157.89 tokens/sec. Stepping up to Qwen3 32B it holds roughly 39.87 tok/s. The full Llama 3.3 70B still runs, at about 18.4 tok/s. For image generation, SDXL runs at 5.35 it/s, and FLUX.1-dev at 1.04 it/s. All 12 workloads fit in 48GB. There is no model in our suite this card has to turn down. NVIDIA RTX 6000 Ada Generation isn't a retail purchase for most people. It's rented by the hour. You can run this exact card on RunPod.
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Qwen3 4B | 241.33 tok/s | 2.9 GB peak176 W62°C1.37 tok/WQ4_K_M | ✓ Measured |
| Llama 3.1 8B | 157.89 tok/s | 4.9 GB peak210 W65°C0.75 tok/WQ4_K_M | ✓ Measured |
| Qwen2.5-Coder 14B | 87.51 tok/s | 8.6 GB peak229 W70°C0.38 tok/WQ4_K_M | ✓ Measured |
| Qwen3 32B | 39.87 tok/s | 18.9 GB peak225 W74°C0.18 tok/WQ4_K_M | ✓ Measured |
| Llama 3.3 70B | 18.4 tok/s | 40 GB peak212 W80°C0.09 tok/WQ4_K_M | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Stable Diffusion XL | 10.7 images/min | 14.9 GB peak298 W84°C5.6 s/img | ✓ Measured |
| Z-Image Turbo | 5.1 images/min | 25.8 GB peak298 W88°C11.8 s/img | ✓ Measured |
| FLUX.1 dev | 2.23 images/min | 36.7 GB peak299 W89°C26.8 s/img | ✓ Measured |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| FLUX.1 Kontext dev | 1.03 images/min | 35.5 GB peak298 W89°C57.9 s/img | ✓ Measured |
| Architecture | Ada Lovelace |
| CUDA cores | 18,176 |
| VRAM | 48GB GDDR6 ECC |
| Memory bus | 384-bit |
| Memory bandwidth | 960 GB/s |
| Boost clock | 2,505 MHz |
| TDP | 300 W |
| Process | 4nm |
| Interface | PCIe 4.0 x16 |
| Release date | 2022-12-03 |
| Launch MSRP | $6,799 |
NVIDIA RTX 6000 Ada Generation scores 23.8/100, #18 of 102. It ran all 12 workloads. Every figure here is our own measurement.
100% = this card, AI & Machine Learning headline metric (AI Score). #1 of 61 desktop cards in this vertical.
| GPU | Relative | % | AI Score |
|---|---|---|---|
| NVIDIA RTX 6000 Ada Generation | 100% | 23.8 | |
| NVIDIA GeForce RTX 5090 | 94% | 22.3 | |
| NVIDIA RTX 5880 Ada Generation | 68% | 16.3 | |
| NVIDIA RTX 5000 Ada Generation | 46% | 11 | |
| NVIDIA GeForce RTX 4090 | 44% | 10.4 |
← All AI & Machine Learning GPU rankings
Whole-job timings, composed from our measured per-model results on this card.
| Workflow | Time | Energy | Basis |
|---|---|---|---|
| 24-frame storyboard | 5.5 min | 25.57 Wh | all 2 stages measured |
| 6-panel comic page | 9.3 min | 43.46 Wh | all 3 stages measured |
| Full codebase review | 11.4 min | 43.65 Wh | measured |
| Character sheet, 12 poses | 12.6 min | 60.12 Wh | all 2 stages measured |
| 20 long-form articles | 25.4 min | 89.53 Wh | measured |
| 40-product photo shoot | 43 min | 211.51 Wh | all 2 stages measured |
| 100-photo restoration batch | 1 h 37 min | 482.35 Wh | measured |
This card is $6,799 to buy. The cheapest listed rate on RunPod is $0.740/hour, but that is the floor: we budget $0.888/hour, a 20% premium, because idle time, storage and unavailable cheap instances all land on the same bill. At that rate buying wins after 7,657 GPU-hours. Below it you are paying for idle silicon.
| How you would use it | GPU-hours a year | Rental cost a year | Time to break even |
|---|---|---|---|
| 2 hours a day, hobby | 730 | $648 | 10.5 years |
| 8 hours a day, working on it | 2,920 | $2,593 | 2.6 years |
| 24/7, always-on agent | 8,760 | $7,779 | 10.5 months |
At hobby usage this card is very unlikely to pay for itself before it is superseded. Rent it. Rental figures include a 20% premium over the cheapest listed rate. Ignores electricity, resale and the fact that a rented card can be a newer one tomorrow.
Cheapest of the RunPod and Vast on-demand rates we see, sampled daily. Spot and interruptible pricing runs lower.