6-panel comic page

Plot the page, draw every panel at FLUX quality, then fix the details in edit.

The pipeline

  1. Plot the page and write dialogue, qwen3-32b
  2. Draw 6 panels, flux-1-dev
  3. Clean up each panel, flux-1-kontext-dev

Fastest card where every stage is a real measurement: NVIDIA B300, 2 min for the whole job.

Every GPU, slowest job to fastest

GPUTotal timeComputeEnergyBasis
NVIDIA B3002 min61 s15.52 Whall 3 stages measured
NVIDIA B1002 min2 minn/aanchored estimate (0/3 stages measured)
NVIDIA GH200 Grace Hopper2.1 min2.1 minn/aanchored estimate (0/3 stages measured)
NVIDIA H100 NVL2.2 min2.2 minn/aanchored estimate (0/3 stages measured)
NVIDIA H800 80GB2.2 min2.2 minn/aanchored estimate (0/3 stages measured)
NVIDIA B2002.3 min1.7 min23.05 Whall 3 stages measured
NVIDIA H100 PCIe2.7 min2.7 minn/aanchored estimate (0/3 stages measured)
NVIDIA H100 80GB HBM33 min2.2 min23.89 Whall 3 stages measured
NVIDIA H2003.2 min2.1 min23.21 Whall 3 stages measured
NVIDIA RTX PRO 6000 Blackwell Server Edition3.6 min3.1 min26.83 Whall 3 stages measured
NVIDIA RTX PRO 6000 Blackwell Max-Q Workstation Edition3.8 min3.8 minn/aanchored estimate (0/3 stages measured)
NVIDIA A800 80GB4.7 min4.7 minn/aanchored estimate (0/3 stages measured)
NVIDIA A100 40GB SXM44.8 min4.7 min1.04 Whanchored estimate (1/3 stages measured)
NVIDIA RTX PRO 6000 Blackwell Workstation Edition5.1 min3 min28.55 Whall 3 stages measured
NVIDIA A100 40GB PCIe5.1 min5.1 minn/aanchored estimate (0/3 stages measured)
NVIDIA RTX PRO 5000 Blackwell5.6 min5 min24.38 Whall 3 stages measured
NVIDIA A100 80GB PCIe5.9 min5 min24.31 Whall 3 stages measured
NVIDIA A100 80GB SXM46 min4.7 min29.8 Whall 3 stages measured
NVIDIA L40S6.5 min5.4 min30.3 Whall 3 stages measured
NVIDIA L409.1 min8.2 min39.9 Whall 3 stages measured
NVIDIA RTX 6000 Ada Generation9.3 min8.8 min43.46 Whall 3 stages measured
NVIDIA RTX PRO 4500 Blackwell11.6 min11.4 min26.25 Whall 3 stages measured
NVIDIA RTX A600012.1 min8.9 min42.76 Whall 3 stages measured
NVIDIA RTX 5000 Ada Generation12.8 min12.7 min35.55 Whall 3 stages measured
NVIDIA RTX 5880 Ada Generation14 min14 minn/aanchored estimate (0/3 stages measured)
AMD Radeon Pro W790017 min17 minn/aanchored estimate (0/3 stages measured)
AMD Radeon Pro W780027.7 min27.7 minn/aanchored estimate (0/3 stages measured)
AMD Radeon Pro W680031.5 min31.5 minn/aanchored estimate (0/3 stages measured)

Cards that can't run this pipeline

Every one of these fails on the same kind of wall, a stage that will not fit in VRAM.

…and 30 more.

How these numbers are built

Each stage time is the quantity of work divided by that card's measured throughput for that model, from our own bench. The pipeline is assumed to run batched, every image, then every clip, so each model loads once. Model load time is added where we recorded it; our text-generation runs don't carry a load measurement yet, so pipelines with a language-model stage are slightly optimistic. Nothing here is a single timed run of the whole pipeline, and we don't present it as one.