60-second AI short film

Script it, generate every keyframe, then render the clips. The full text-to-video pipeline end to end.

The pipeline

  1. Write the script and shot list, qwen3-32b
  2. Generate 15 keyframes, stable-diffusion-xl
  3. Render 15 clips (~4s each), ltx-video

Fastest card where every stage is a real measurement: NVIDIA B200, 1.9 min for the whole job.

Every GPU, slowest job to fastest

GPUTotal timeComputeEnergyBasis
NVIDIA B1001.7 min1.7 minn/aanchored estimate (0/3 stages measured)
NVIDIA B2001.9 min85 s16.03 Whall 3 stages measured
NVIDIA GH200 Grace Hopper2 min2 minn/aanchored estimate (0/3 stages measured)
NVIDIA H100 NVL2 min2 minn/aanchored estimate (0/3 stages measured)
NVIDIA H800 80GB2 min2 minn/aanchored estimate (0/3 stages measured)
NVIDIA B3002.2 min87 s14.29 Whall 3 stages measured
NVIDIA H100 PCIe2.4 min2.4 minn/aanchored estimate (0/3 stages measured)
NVIDIA H100 80GB HBM32.5 min2 min17.46 Whall 3 stages measured
NVIDIA RTX PRO 6000 Blackwell Server Edition2.7 min2.4 min19.22 Whall 3 stages measured
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition2.8 min2.8 minn/aanchored estimate (0/3 stages measured)
NVIDIA H2003.9 min2 min18.03 Whall 3 stages measured
NVIDIA A800 80GB4 min4 minn/aanchored estimate (0/3 stages measured)
NVIDIA A100 40GB SXM44 min3.9 min6.39 Whanchored estimate (2/3 stages measured)
NVIDIA A100 40GB PCIe4.2 min4.2 minn/aanchored estimate (0/3 stages measured)
NVIDIA RTX PRO 6000 Blackwell Workstation Edition4.5 min2.2 min19.26 Whall 3 stages measured
NVIDIA A100 80GB PCIe4.6 min4.1 min18.6 Whall 3 stages measured
NVIDIA A100 80GB SXM44.7 min4 min23.61 Whall 3 stages measured
NVIDIA L40S4.9 min4.3 min22.52 Whall 3 stages measured
NVIDIA RTX PRO 5000 Blackwell6.7 min6.4 min20.33 Whall 3 stages measured
NVIDIA GeForce RTX 40907.9 min6.4 min30.54 Whall 3 stages measured
NVIDIA RTX 5880 Ada Generation8.6 min8.6 minn/aanchored estimate (0/3 stages measured)
NVIDIA RTX 5000 Ada Generation9.5 min9.3 min27.06 Whall 3 stages measured
NVIDIA GeForce RTX 3090 Ti9.9 min9.7 min50.99 Whall 3 stages measured
NVIDIA RTX PRO 4000 Blackwell10.2 min10.1 min20.63 Whall 3 stages measured
NVIDIA RTX PRO 4500 Blackwell10.6 min8.3 min18.99 Whall 3 stages measured
NVIDIA RTX A500011.8 min11.5 min36.17 Whall 3 stages measured
NVIDIA L4013 min11.1 min36.37 Whall 3 stages measured
NVIDIA RTX A550013.1 min13.1 minn/aanchored estimate (0/3 stages measured)
NVIDIA RTX 4500 Ada Generation15.1 min15.1 minn/aanchored estimate (0/3 stages measured)
NVIDIA GeForce RTX 309015.2 min13.6 min57.1 Whall 3 stages measured
AMD Radeon RX 7900 XTX16.1 min16.1 minn/aanchored estimate (0/3 stages measured)
AMD Radeon Pro W790018.6 min18.6 minn/aanchored estimate (0/3 stages measured)
AMD Radeon RX 7900 XT20.7 min20.7 minn/aanchored estimate (0/3 stages measured)
AMD Radeon Pro W780024.3 min24.3 minn/aanchored estimate (0/3 stages measured)
AMD Radeon Pro W680028.8 min28.8 minn/aanchored estimate (0/3 stages measured)

Cards that can't run this pipeline

Every one of these fails on the same kind of wall, a stage that will not fit in VRAM.

…and 18 more.

How these numbers are built

Each stage time is the quantity of work divided by that card's measured throughput for that model, from our own bench. The pipeline is assumed to run batched, every image, then every clip, so each model loads once. Model load time is added where we recorded it; our text-generation runs don't carry a load measurement yet, so pipelines with a language-model stage are slightly optimistic. Nothing here is a single timed run of the whole pipeline, and we don't present it as one.