Z-Image Turbo · 44 cards measured first-party · Updated October 2026

How Fast Is Z-Image Turbo on Each GPU?

Z-Image Turbo is a modern distilled image model, 8 steps instead of 30, which sounds like it should be trivial. It isn't: it wants ~13GB minimum and ~23GB to run clean, so it gates hardware that handles SDXL without complaint.

Benchmarked weights: Tongyi-MAI/Z-Image-Turbo

Fastest we measured
NVIDIA B300

NVIDIA B300

39.67 images/min on Z-Image Turbo, the ceiling. Measured on our bench. 288GB of VRAM, $40,000 at launch.

Pros
  • 39.67 images/min on Z-Image Turbo
  • 288GB, clears the Z-Image Turbo floor
  • Rentable by the hour rather than bought
Cons
  • 1400W board rating
  • Datacenter or workstation hardware, not a retail purchase

Best for: Z-Image Turbo work where you want the ceiling gone rather than the cheapest entry.

Best consumer card
NVIDIA GeForce RTX 5090

NVIDIA GeForce RTX 5090

10.72 images/min on Z-Image Turbo, fastest card you can buy at retail. Measured on our bench. 32GB of VRAM, $1,999 at launch.

Pros
  • 10.72 images/min on Z-Image Turbo
  • 32GB, clears the Z-Image Turbo floor
Cons
  • 575W board rating
Cheapest card that runs it
GeForce RTX 5060 Ti

GeForce RTX 5060 Ti

1.56 images/min on Z-Image Turbo, lowest launch price that still fits. Measured on our bench. 16GB of VRAM, $429 at launch.

Pros
  • 1.56 images/min on Z-Image Turbo
  • 16GB, clears the Z-Image Turbo floor
Cons
  • 180W board rating
Best value
NVIDIA GeForce RTX 4090

NVIDIA GeForce RTX 4090

7.24 images/min on Z-Image Turbo, most speed per dollar. Measured on our bench. 24GB of VRAM, $1,599 at launch. That is 4.53 images/min per $1,000 of launch price.

Pros
  • 7.24 images/min on Z-Image Turbo
  • 24GB, clears the Z-Image Turbo floor
Cons
  • 450W board rating
39.67images/min
Fastest: NVIDIA B300
measured
45
Cards that run Z-Image Turbo
of 76 we have data for
31
Cards that can't run it at all
published as hard gates, not omissions
8717%
Fastest vs slowest that fits
39.67 vs 0.45 images/min

Compute-bound like all diffusion, but the step count changes the character of the benchmark. Fewer, heavier steps means less opportunity to hide latency, and the gap between architecture generations widens rather than narrows.

Z-Image Turbo: speed on every GPU we have data for

NVIDIA B300
39.67 images/min
NVIDIA B200
32.33 images/min
NVIDIA B100
25.88 images/min
NVIDIA GH200 Grace Hopper
23.18 images/min
NVIDIA H200
23.18 images/min
NVIDIA H100 NVL
22.12 images/min
NVIDIA H100 80GB HBM3
22.12 images/min
NVIDIA H800 80GB
22.12 images/min
NVIDIA H100 PCIe
19.12 images/min
NVIDIA RTX PRO 6000 Blackwell Workstation Edition
15.6 images/min
NVIDIA RTX PRO 6000 Blackwell Server Edition
14.78 images/min
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition
12.15 images/min
NVIDIA GeForce RTX 5090
10.72 images/min
NVIDIA A100 40GB SXM4
9.97 images/min
NVIDIA A100 80GB SXM4
9.97 images/min

Top 15 shown; 30 more cards in the full table below.

Single stream, batch size 1. 34 of the 56 cards on this page were measured first-party by us; the rest are anchored estimates against those measurements and are labelled in the table below.

Efficiency: images/min per 100W drawn

NVIDIA B300
4.26 images/min / 100W
NVIDIA B200
3.6 images/min / 100W
NVIDIA H200
3.49 images/min / 100W
NVIDIA L4
3.45 images/min / 100W
NVIDIA H100 80GB HBM3
3.35 images/min / 100W
NVIDIA RTX PRO 4500 Blackwell
3.23 images/min / 100W
NVIDIA A100 80GB PCIe
3.18 images/min / 100W
NVIDIA RTX PRO 4000 Blackwell
3.05 images/min / 100W
NVIDIA RTX PRO 5000 Blackwell
3.03 images/min / 100W
NVIDIA RTX PRO 6000 Blackwell Server Edition
2.78 images/min / 100W
NVIDIA RTX PRO 6000 Blackwell Workstation Edition
2.67 images/min / 100W
NVIDIA A100 80GB SXM4
2.61 images/min / 100W
NVIDIA RTX 5000 Ada Generation
2.59 images/min / 100W
NVIDIA L40S
2.53 images/min / 100W
NVIDIA A10G
2.32 images/min / 100W

Top 15 shown; 13 more cards in the full table below.

Power is the average pulled during the run, sampled at 1Hz. The fastest card is often not the one here, and for anything left running this is the number that shows up on the bill.

Value: images/min per $1,000 of MSRP

NVIDIA GeForce RTX 5090
5.37 images/min / $1k
NVIDIA GeForce RTX 4090
4.53 images/min / $1k
GeForce RTX 5060 Ti
3.64 images/min / $1k
GeForce RTX 5080
3.56 images/min / $1k
GeForce RTX 5070 Ti
3.24 images/min / $1k
NVIDIA RTX PRO 4000 Blackwell
2.95 images/min / $1k
NVIDIA GeForce RTX 4070 Ti Super
2.88 images/min / $1k
NVIDIA GeForce RTX 4060 Ti 16GB
2.83 images/min / $1k
NVIDIA GeForce RTX 4080
2.56 images/min / $1k
NVIDIA GeForce RTX 3090
2.55 images/min / $1k
NVIDIA RTX PRO 4500 Blackwell
2.48 images/min / $1k
GeForce RTX 4080 Super
2.3 images/min / $1k
NVIDIA GeForce RTX 3090 Ti
2.18 images/min / $1k
NVIDIA RTX PRO 5000 Blackwell
2.02 images/min / $1k
NVIDIA RTX PRO 6000 Blackwell Workstation Edition
1.82 images/min / $1k

Top 15 shown; 17 more cards in the full table below.

Launch price, not street price, so it ages. A speed leaderboard always crowns the most expensive card; this is the counterweight.

Won't fit, Z-Image Turbo gates these cards outright

NVIDIA GeForce RTX 306012GB
NVIDIA GeForce RTX 3080 Ti12GB
NVIDIA GeForce RTX 4070 Super12GB
NVIDIA GeForce RTX 4070 Ti12GB
NVIDIA GeForce RTX 407012GB
GeForce RTX 507012GB
NVIDIA TITAN V12GB
NVIDIA TITAN X (Pascal)12GB
NVIDIA TITAN Xp12GB
GeForce GTX 1080 Ti11GB
NVIDIA GeForce RTX 2080 Ti Founders Edition11GB
NVIDIA GeForce RTX 308010GB
NVIDIA GeForce GTX 1070 Ti8GB
NVIDIA GeForce GTX 10808GB
NVIDIA GeForce RTX 2060 Super8GB
NVIDIA GeForce RTX 2070 SUPER8GB
NVIDIA GeForce RTX 20708GB
NVIDIA GeForce RTX 2080 Super8GB
NVIDIA GeForce RTX 2080 Founders Edition8GB
NVIDIA GeForce RTX 30508GB
NVIDIA GeForce RTX 3060 Ti8GB
NVIDIA GeForce RTX 3070 Ti8GB
NVIDIA GeForce RTX 3070 Founders Edition8GB
GeForce RTX 40608GB
NVIDIA GeForce RTX 50508GB
NVIDIA GeForce RTX 50608GB
NVIDIA GeForce GTX 1660 Super6GB
NVIDIA GeForce GTX 1660 Ti6GB
NVIDIA GeForce GTX 16606GB
NVIDIA GeForce RTX 20606GB
NVIDIA RTX A20006GB
GPUVRAMWhy it fails
NVIDIA GeForce RTX 306012GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 3080 Ti12GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 4070 Super12GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 4070 Ti12GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 407012GBNeeds ~13GB VRAM
GeForce RTX 507012GBNeeds ~13GB VRAM
NVIDIA TITAN V12GBNeeds ~13GB VRAM
NVIDIA TITAN X (Pascal)12GBNeeds ~13GB VRAM
NVIDIA TITAN Xp12GBNeeds ~13GB VRAM
GeForce GTX 1080 Ti11GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 2080 Ti Founders Edition11GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 308010GBNeeds ~13GB VRAM
NVIDIA GeForce GTX 1070 Ti8GBNeeds ~13GB VRAM
NVIDIA GeForce GTX 10808GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 2060 Super8GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 2070 SUPER8GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 20708GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 2080 Super8GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 2080 Founders Edition8GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 30508GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 3060 Ti8GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 3070 Ti8GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 3070 Founders Edition8GBNeeds ~13GB VRAM
GeForce RTX 40608GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 50508GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 50608GBNeeds ~13GB VRAM
NVIDIA GeForce GTX 1660 Super6GBNeeds ~13GB VRAM
NVIDIA GeForce GTX 1660 Ti6GBNeeds ~13GB VRAM
NVIDIA GeForce GTX 16606GBNeeds ~13GB VRAM
NVIDIA GeForce RTX 20606GBNeeds ~13GB VRAM
NVIDIA RTX A20006GBNeeds ~13GB VRAM

No driver update fixes a VRAM ceiling.

Full Z-Image Turbo leaderboard, every card that runs it

NVIDIA B30039.67 images/min
NVIDIA B20032.33 images/min
NVIDIA B10025.88 images/min
NVIDIA GH200 Grace Hopper23.18 images/min
NVIDIA H20023.18 images/min
NVIDIA H100 NVL22.12 images/min
NVIDIA H100 80GB HBM322.12 images/min
NVIDIA H800 80GB22.12 images/min
NVIDIA H100 PCIe19.12 images/min
NVIDIA RTX PRO 6000 Blackwell Workstation Edition15.6 images/min
NVIDIA RTX PRO 6000 Blackwell Server Edition14.78 images/min
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition12.15 images/min
NVIDIA GeForce RTX 509010.72 images/min
NVIDIA A100 40GB SXM49.97 images/min
NVIDIA A100 80GB SXM49.97 images/min
NVIDIA A800 80GB9.97 images/min
NVIDIA A100 40GB PCIe9.3 images/min
NVIDIA A100 80GB PCIe9.3 images/min
NVIDIA RTX PRO 5000 Blackwell9.07 images/min
NVIDIA L40S8.32 images/min
NVIDIA RTX A55007.4 images/min
NVIDIA GeForce RTX 40907.24 images/min
NVIDIA RTX PRO 4500 Blackwell6.45 images/min
NVIDIA RTX 5000 Ada Generation5.92 images/min
NVIDIA L405.62 images/min
NVIDIA RTX A60005.4 images/min
NVIDIA RTX 6000 Ada Generation5.1 images/min
NVIDIA RTX PRO 4000 Blackwell4.42 images/min
NVIDIA GeForce RTX 3090 Ti4.35 images/min
NVIDIA RTX A50004.05 images/min
NVIDIA GeForce RTX 30903.83 images/min
GeForce RTX 50803.56 images/min
NVIDIA A10G3.45 images/min
NVIDIA RTX 5880 Ada Generation3.24 images/min
NVIDIA A403.23 images/min
NVIDIA GeForce RTX 40803.07 images/min
NVIDIA L42.48 images/min
GeForce RTX 5070 Ti2.43 images/min
NVIDIA GeForce RTX 4070 Ti Super2.3 images/min
GeForce RTX 4080 Super2.3 images/min
NVIDIA RTX 4500 Ada Generation1.73 images/min
GeForce RTX 5060 Ti1.56 images/min
NVIDIA Quadro RTX 50001.5 images/min
NVIDIA GeForce RTX 4060 Ti 16GB1.41 images/min
NVIDIA Quadro RTX 80000.45 images/min
GPUResultVRAMSource
NVIDIA B30039.67 images/min288GBMeasured
NVIDIA B20032.33 images/min192GBMeasured
NVIDIA B10025.88 images/min192GBEstimated
NVIDIA GH200 Grace Hopper23.18 images/min141GBEstimated
NVIDIA H20023.18 images/min141GBMeasured
NVIDIA H100 NVL22.12 images/min94GBEstimated
NVIDIA H100 80GB HBM322.12 images/min80GBMeasured
NVIDIA H800 80GB22.12 images/min80GBEstimated
NVIDIA H100 PCIe19.12 images/min80GBEstimated
NVIDIA RTX PRO 6000 Blackwell Workstation Edition15.6 images/min96GBMeasured
NVIDIA RTX PRO 6000 Blackwell Server Edition14.78 images/min96GBMeasured
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition12.15 images/min96GBEstimated
NVIDIA GeForce RTX 509010.72 images/min32GBMeasured
NVIDIA A100 40GB SXM49.97 images/min40GBEstimated
NVIDIA A100 80GB SXM49.97 images/min80GBMeasured
NVIDIA A800 80GB9.97 images/min80GBEstimated
NVIDIA A100 40GB PCIe9.3 images/min40GBEstimated
NVIDIA A100 80GB PCIe9.3 images/min80GBMeasured
NVIDIA RTX PRO 5000 Blackwell9.07 images/min48GBMeasured
NVIDIA L40S8.32 images/min48GBMeasured
NVIDIA RTX A55007.4 images/min24GBEstimated
NVIDIA GeForce RTX 40907.24 images/min24GBMeasured
NVIDIA RTX PRO 4500 Blackwell6.45 images/min32GBMeasured
NVIDIA RTX 5000 Ada Generation5.92 images/min32GBMeasured
NVIDIA L405.62 images/min48GBMeasured
NVIDIA RTX A60005.4 images/min48GBMeasured
NVIDIA RTX 6000 Ada Generation5.1 images/min48GBMeasured
NVIDIA RTX PRO 4000 Blackwell4.42 images/min24GBMeasured
NVIDIA GeForce RTX 3090 Ti4.35 images/min24GBMeasured
NVIDIA RTX A50004.05 images/min24GBMeasured
NVIDIA GeForce RTX 30903.83 images/min24GBMeasured
GeForce RTX 50803.56 images/min16GBMeasured
NVIDIA A10G3.45 images/min24GBMeasured
NVIDIA RTX 5880 Ada Generation3.24 images/min48GBEstimated
NVIDIA A403.23 images/min48GBMeasured
NVIDIA GeForce RTX 40803.07 images/min16GBMeasured
NVIDIA L42.48 images/min24GBMeasured
GeForce RTX 5070 Ti2.43 images/min16GBMeasured
NVIDIA GeForce RTX 4070 Ti Super2.3 images/min16GBMeasured
GeForce RTX 4080 Super2.3 images/min16GBMeasured
NVIDIA RTX 4500 Ada Generation1.73 images/min24GBEstimated
GeForce RTX 5060 Ti1.56 images/min16GBMeasured
NVIDIA Quadro RTX 50001.5 images/min16GBEstimated
NVIDIA GeForce RTX 4060 Ti 16GB1.41 images/min16GBMeasured
NVIDIA Quadro RTX 80000.45 images/min48GBMeasured

Tap any column to sort. Measured = we rented and ran this card ourselves. Estimated = interpolated against our measured anchors, never blended silently.

Because this workload is tensor-compute bound, the ranking tracks architecture generation and tensor throughput rather than memory bandwidth, the reverse of our LLM charts. The same two cards can swap places entirely depending on which of these pages you're reading. That's the reason we run twelve workloads instead of publishing one score. A GPU isn't fast or slow. It's fast at some things and gated out of others, and which of those matters depends entirely on what you're actually going to run.

Our verdict

NVIDIA B300 tops our Z-Image Turbo leaderboard at 39.67 images/min (measured), 8717% of the way clear of the slowest card that still fits. But the number that decides most purchases isn't on the chart. It's the 31 cards that can't run Z-Image Turbo at all. This is a compute workload: buy architecture generation, not raw VRAM, as long as you clear the floor first.

FAQ

What is the fastest GPU for Z-Image Turbo?
NVIDIA B300, at 39.67 images/min on our bench, a first-party measurement. It carries 288GB of VRAM. Of the 76 cards we have Z-Image Turbo data for, 45 can run it at all.
How much VRAM do I need for Z-Image Turbo?
~13GB minimum, ~23GB for the full path. Cards that offload here take a brutal penalty rather than a graceful one.
Why does the Z-Image Turbo ranking look different from your other benchmarks?
Because this workload is tensor-compute bound, the ranking tracks architecture generation and tensor throughput rather than memory bandwidth, the reverse of our LLM charts. The same two cards can swap places entirely depending on which of these pages you're reading. That's why we publish twelve separate workloads rather than one blended score, the ordering genuinely changes depending on the job.
Are these Z-Image Turbo numbers measured or estimated?
Both, and every row says which. 44 of the 76 cards here were rented and run by us on the same harness. The remainder are anchored estimates interpolated per workload against those measurements. We never blend the two silently, if a row says Estimated, we have not run that card.
Can I rent a GPU to run Z-Image Turbo instead of buying one?
Yes, and for the cards at the top of this leaderboard it's the only realistic option, most of them have no retail channel at all. It's also how we got these numbers: we rented the hardware by the hour rather than buying it. That's worth considering before you spend on a card to find out whether it's fast enough.
Why publish cards that can't run Z-Image Turbo?
Because it's the most useful thing we know. A card that can't load a model doesn't run it slowly, it doesn't run it. Most benchmark sites leave that as a blank cell or quietly drop to a smaller quantisation to produce a number. We publish it as a hard gate and score it zero, because 'this card cannot do the thing you want' is the answer to the question you were actually asking.

How we test

Every ranking on this page comes from our own benchmark runs, not vendor claims. Cards marked Measured were rented and run by us; cards marked Estimated are interpolated per workload against those measured anchors and are labelled on every row, we never blend the two silently. LLMs run on llama.cpp (llama-bench) at Q4_K_M with -p 512 -n 128. Diffusion and video run on diffusers/ComfyUI at BF16, with SDXL at FP16. Each workload gets a warmup pass plus multiple timed runs (5 for small LLMs, 3 for large models and images, 2 for video); we publish the mean as the result and the minimum as the 1% low. Run-to-run variance is under 0.5%. Telemetry, power, temperature, utilisation, clocks, peak VRAM, is sampled at 1 Hz for the duration of every run. Where a model exceeds a card's VRAM we publish a hard won't-fit result rather than quietly dropping to a smaller quantisation. A card that can't run a model scores zero on it. Silently swapping precision to make a number appear would make every number on this site meaningless. All figures are single-GPU, single-stream, batch-size-1. That is the honest way to measure what one card does for one user, and it is deliberately not how a datacenter serves a model. Vendor and MLPerf figures use large batches across many GPUs and will be far higher. Neither is wrong, they answer different questions. Ours answers 'what will this card do for me'.