Character sheet, 12 poses

One hero render, then twelve consistent variations, the workflow behind every game and animation pitch.

The pipeline

  1. Render the hero image, flux-1-dev
  2. Generate 12 consistent poses, flux-1-kontext-dev

Fastest card where every stage is a real measurement: NVIDIA B300, 2.2 min for the whole job.

Every GPU, slowest job to fastest

GPUTotal timeComputeEnergyBasis
NVIDIA B3002.2 min73 s20.73 Whall 2 stages measured
NVIDIA B1002.7 min2.7 minn/aanchored estimate (0/2 stages measured)
NVIDIA GH200 Grace Hopper2.8 min2.8 minn/aanchored estimate (0/2 stages measured)
NVIDIA B2002.9 min2.2 min32.04 Whall 2 stages measured
NVIDIA H100 NVL2.9 min2.9 minn/aanchored estimate (0/2 stages measured)
NVIDIA H800 80GB2.9 min2.9 minn/aanchored estimate (0/2 stages measured)
NVIDIA H100 PCIe3.4 min3.4 minn/aanchored estimate (0/2 stages measured)
NVIDIA H100 80GB HBM33.7 min2.9 min33.26 Whall 2 stages measured
NVIDIA H2003.9 min2.8 min32.58 Whall 2 stages measured
NVIDIA RTX PRO 6000 Blackwell Server Edition4.6 min4.1 min37.55 Whall 2 stages measured
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition5.2 min5.2 minn/aanchored estimate (0/2 stages measured)
NVIDIA RTX PRO 6000 Blackwell Workstation Edition6.1 min4 min40.17 Whall 2 stages measured
NVIDIA A100 40GB SXM46.3 min6.3 minn/aanchored estimate (0/2 stages measured)
NVIDIA A800 80GB6.3 min6.3 minn/aanchored estimate (0/2 stages measured)
NVIDIA A100 40GB PCIe6.8 min6.8 minn/aanchored estimate (0/2 stages measured)
NVIDIA RTX PRO 5000 Blackwell7.6 min6.9 min34.5 Whall 2 stages measured
NVIDIA A100 80GB SXM47.6 min6.3 min41.23 Whall 2 stages measured
NVIDIA A100 80GB PCIe7.7 min6.8 min33.46 Whall 2 stages measured
NVIDIA L40S8.4 min7.3 min41.61 Whall 2 stages measured
NVIDIA L4012.1 min11.2 min55.81 Whall 2 stages measured
NVIDIA RTX 6000 Ada Generation12.6 min12.1 min60.12 Whall 2 stages measured
NVIDIA RTX PRO 4500 Blackwell14.6 min14.4 min35.78 Whall 2 stages measured
NVIDIA RTX A600015.2 min11.9 min58.73 Whall 2 stages measured
NVIDIA RTX 5000 Ada Generation16 min15.9 min46.88 Whall 2 stages measured
NVIDIA RTX 5880 Ada Generation18.3 min18.3 minn/aanchored estimate (0/2 stages measured)
AMD Radeon Pro W790024.2 min24.2 minn/aanchored estimate (0/2 stages measured)
AMD Radeon Pro W780036.6 min36.6 minn/aanchored estimate (0/2 stages measured)
AMD Radeon Pro W680039.4 min39.4 minn/aanchored estimate (0/2 stages measured)

Cards that can't run this pipeline

Every one of these fails on the same kind of wall, a stage that will not fit in VRAM.

…and 31 more.

How these numbers are built

Each stage time is the quantity of work divided by that card's measured throughput for that model, from our own bench. The pipeline is assumed to run batched, every image, then every clip, so each model loads once. Model load time is added where we recorded it; our text-generation runs don't carry a load measurement yet, so pipelines with a language-model stage are slightly optimistic. Nothing here is a single timed run of the whole pipeline, and we don't present it as one.