12GB · AI Score 2.9/100 · anchored estimate vs 51 measured cards
2.9 AI Score Includes estimates
We have not run NVIDIA TITAN V on our bench. These figures are anchored estimates, interpolated per workload against the 51 GPUs we did measure. On Llama 3.1 8B (Q4_K_M) NVIDIA TITAN V should deliver about 104.1 tokens/sec. Llama 3.3 70B does not fit. It needs roughly 42GB and this card has 12GB. For image generation, SDXL should run near 2.22 it/s, while FLUX.1-dev won't fit at BF16 (needs ~26GB). 8 of the 12 workloads won't fit on 12GB at the tested precision, Qwen3 32B, Llama 3.3 70B, Z-Image Turbo, FLUX.1-dev and others. We publish those as hard gates rather than quietly dropping to a smaller quant.
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Qwen3 4B | 154.8 tok/s | estimated | Est. |
| Llama 3.1 8B | 104.1 tok/s | estimated | Est. |
| Qwen2.5-Coder 14B | 54.3 tok/s | estimated | Est. |
| Qwen3 32B | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Llama 3.3 70B | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| Stable Diffusion XL | 4.44 images/min | estimated | Est. |
| Z-Image Turbo | ✕ Won't fit | VRAM-gated at this precision | Est. |
| FLUX.1 dev | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| FLUX.1 Kontext dev | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Qwen-Image-Edit | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Workload | Result | Telemetry | Data |
|---|---|---|---|
| LTX-Video (distilled) | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Wan 2.2 5B (720p) | ✕ Won't fit | VRAM-gated at this precision | Est. |
| Architecture | Volta (GV100) |
| CUDA cores | 5,120 |
| VRAM | 12GB HBM2 |
| Memory bus | 3072-bit |
| Memory bandwidth | 652.8 GB/s |
| Boost clock | 1,455 MHz |
| TDP | 250 W |
| Process | 12nm |
| Interface | PCIe 3.0 x16 |
| Release date | 2017-12-07 |
| Launch MSRP | $2,999 |
NVIDIA TITAN V scores 2.9/100, #71 of 102. It ran 4 of 12; 8 exceeded its 12GB. Figures are anchored estimates, not measurements, we flag that on every row.
100% = this card, AI & Machine Learning headline metric (AI Score). #32 of 61 desktop cards in this vertical.
| GPU | Relative | % | AI Score |
|---|---|---|---|
| NVIDIA GeForce RTX 4070 Super | 107% | 3.1 | |
| NVIDIA GeForce RTX 4070 | 107% | 3.1 | |
| NVIDIA GeForce RTX 4070 Ti | 103% | 3 | |
| Intel Arc A770 Limited Edition | 100% | 2.9 | |
| NVIDIA TITAN V | 100% | 2.9 | |
| GeForce GTX 1080 Ti | 90% | 2.6 | |
| NVIDIA GeForce RTX 3060 | 90% | 2.6 | |
| NVIDIA TITAN Xp | 90% | 2.6 | |
| AMD Radeon RX 7700 XT | 86% | 2.5 |
Same card, other workloads: NVIDIA TITAN V Gaming benchmarks
← All AI & Machine Learning GPU rankings
Whole-job timings, composed from our measured per-model results on this card.
| Workflow | Time | Energy | Basis |
|---|---|---|---|
| Full codebase review | 18.4 min | n/a | estimate, 0 of 1 stage measured |
Can't run: 60-second AI short film (needs Qwen3 32B), 60-second AI short film, narrated (needs Qwen3 32B), 10 short social clips (needs Qwen3 32B), 40-product photo shoot (needs FLUX.1 Kontext dev), 6-panel comic page (needs Qwen3 32B), 20 long-form articles (needs Llama 3.3 70B), Character sheet, 12 poses (needs FLUX.1 dev), 100-photo restoration batch (needs FLUX.1 Kontext dev), 24-frame storyboard (needs Qwen3 32B), 100-photo restore and enlarge (needs FLUX.1 Kontext dev).