24-frame storyboard

Break a script into beats and draw the whole board, fast turbo diffusion, so it lands in minutes.

The pipeline

  1. Break the script into 24 beats, qwen3-32b
  2. Draw 24 storyboard frames, z-image-turbo

Fastest card where every stage is a real measurement: NVIDIA B200, 89 s for the whole job.

Every GPU, slowest job to fastest

GPUTotal timeComputeEnergyBasis
NVIDIA B10074 s74 sn/aanchored estimate (0/2 stages measured)
NVIDIA GH200 Grace Hopper80 s80 sn/aanchored estimate (0/2 stages measured)
NVIDIA H100 NVL81 s81 sn/aanchored estimate (0/2 stages measured)
NVIDIA H800 80GB84 s84 sn/aanchored estimate (0/2 stages measured)
NVIDIA B20089 s62 s12.85 Whall 2 stages measured
NVIDIA B3001.5 min53 s11 Whall 2 stages measured
NVIDIA H100 PCIe1.8 min1.8 minn/aanchored estimate (0/2 stages measured)
NVIDIA H100 80GB HBM31.8 min84 s13.08 Whall 2 stages measured
NVIDIA RTX PRO 6000 Blackwell Server Edition2.3 min2 min15.7 Whall 2 stages measured
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition2.3 min2.3 minn/aanchored estimate (0/2 stages measured)
NVIDIA A800 80GB2.9 min2.9 minn/aanchored estimate (0/2 stages measured)
NVIDIA A100 40GB SXM43 min2.9 min2.08 Whanchored estimate (1/2 stages measured)
NVIDIA A100 40GB PCIe3.2 min3.2 minn/aanchored estimate (0/2 stages measured)
NVIDIA RTX PRO 5000 Blackwell3.4 min3.1 min14.39 Whall 2 stages measured
NVIDIA RTX PRO 6000 Blackwell Workstation Edition3.6 min1.9 min16.08 Whall 2 stages measured
NVIDIA A100 80GB PCIe3.7 min3.1 min13.93 Whall 2 stages measured
NVIDIA A100 80GB SXM43.7 min2.9 min16.68 Whall 2 stages measured
NVIDIA H2003.9 min80 s12.08 Whall 2 stages measured
NVIDIA L40S4.1 min3.6 min18.66 Whall 2 stages measured
NVIDIA GeForce RTX 40905 min3.8 min24.32 Whall 2 stages measured
NVIDIA RTX 5000 Ada Generation5.1 min5 min18.41 Whall 2 stages measured
NVIDIA RTX A55005.3 min5.3 minn/aanchored estimate (0/2 stages measured)
NVIDIA RTX 6000 Ada Generation5.5 min5.3 min25.57 Whall 2 stages measured
NVIDIA RTX 5880 Ada Generation6.2 min6.2 minn/aanchored estimate (0/2 stages measured)
NVIDIA RTX PRO 4500 Blackwell6.2 min4.3 min13.21 Whall 2 stages measured
NVIDIA GeForce RTX 3090 Ti6.3 min6.1 min37.71 Whall 2 stages measured
NVIDIA RTX PRO 4000 Blackwell6.4 min6.3 min14.71 Whall 2 stages measured
NVIDIA RTX A50007 min6.7 min24.57 Whall 2 stages measured
NVIDIA L407 min5 min23.31 Whall 2 stages measured
NVIDIA RTX A60007.1 min5.2 min23.42 Whall 2 stages measured
NVIDIA GeForce RTX 30908.3 min6.9 min37.34 Whall 2 stages measured
NVIDIA A10G9 min8 min19.54 Whall 2 stages measured
AMD Radeon Pro W790010.1 min10.1 minn/aanchored estimate (0/2 stages measured)
NVIDIA A4010.2 min8.3 min39.46 Whall 2 stages measured
AMD Radeon RX 7900 XTX10.3 min10.3 minn/aanchored estimate (0/2 stages measured)
AMD Radeon RX 7900 XT10.6 min10.6 minn/aanchored estimate (0/2 stages measured)
NVIDIA L412.2 min11.6 min13.69 Whall 2 stages measured
NVIDIA GeForce RTX 509012.5 min2.6 min21.46 Whall 2 stages measured
AMD Radeon Pro W780013.5 min13.5 minn/aanchored estimate (0/2 stages measured)
NVIDIA RTX 4500 Ada Generation15.2 min15.2 minn/aanchored estimate (0/2 stages measured)
AMD Radeon Pro W680020.1 min20.1 minn/aanchored estimate (0/2 stages measured)
NVIDIA Quadro RTX 800054.9 min54.4 min223 Whall 2 stages measured

Cards that can't run this pipeline

Every one of these fails on the same kind of wall, a stage that will not fit in VRAM.

…and 18 more.

How these numbers are built

Each stage time is the quantity of work divided by that card's measured throughput for that model, from our own bench. The pipeline is assumed to run batched, every image, then every clip, so each model loads once. Model load time is added where we recorded it; our text-generation runs don't carry a load measurement yet, so pipelines with a language-model stage are slightly optimistic. Nothing here is a single timed run of the whole pipeline, and we don't present it as one.