TRELLIS Image-to-3D · 6 GPUs measured first-party · image-to-3D · Updated October 2026
TRELLIS is Microsoft's first-generation image-to-3D model: drop in a picture, get a textured 3D asset. We measured it on 6 GPUs all the way to a finished GLB file: 167 assets an hour on an H100 (21.5 seconds each), and the 3D generation is only 3-4 seconds of that. The rest is the GLB export, which simplifies the mesh and bakes its textures.
Benchmarked weights: microsoft/TRELLIS-image-large

167.2 assets/hour on TRELLIS Image-to-3D, the ceiling. Measured on our bench. 80GB of VRAM, $30,000 at launch.
What GPU Do You Need for TRELLIS Image-to-3D?, assets/hour by GPU
Measured on our own bench. A card absent from this chart has not been run on this model yet, or cannot fit it.
Value: assets/hour per $1,000 of MSRP
Launch price, not street price, so it ages. A speed leaderboard always crowns the most expensive card; this is the counterweight.
TRELLIS Image-to-3D. Measured 3D generation speed by GPU
| GPU | Assets/hour | s per asset | Peak VRAM | Avg power |
|---|---|---|---|---|
| NVIDIA H100 80GB HBM3 | 167.2 | 21.54 | 11.7GB | — |
| NVIDIA H200 | 134.6 | 26.74 | 11.7GB | — |
| NVIDIA L40S | 134.5 | 26.76 | 11.7GB | — |
| NVIDIA A100 80GB SXM4 | 116.4 | 30.93 | 11.7GB | — |
| NVIDIA A10G | 94 | 38.31 | 11.7GB | — |
| NVIDIA L4 | 68.8 | 52.35 | 11.7GB | — |
Where v1 stands now: the volume tool. TRELLIS.2 has superseded this model on quality, and for hero assets that's the page to read. Measured honestly, to a finished GLB on both, v1 is about twice as fast on the same card and needs half the memory. For bulk props, background clutter and candidate shapes to pick from, that's the trade worth making.
Where the time goes. The 3D generation is the quick part: 3.3-4.2 seconds per asset on an H100, 5.3-6.9 on an L40S, 12.8-24.9 on an L4. The GLB export (mesh simplification plus a texture bake that runs on the GPU) takes 70-85% of every asset's time, and it grows with detail: 9 seconds for a pistol on an H100, 45 for a leafy tree. So the card matters less than the chart suggests for simple props, and a lot for detailed ones. If you only need a quick look, TRELLIS's splat preview skips the export entirely.
About TRELLIS Image-to-3D. TRELLIS Image-to-3D: from microsoft, on Hugging Face since December 2024, MIT licence. 2,209,080 downloads in the last 30 days and 19 community quantizations.
How it compares. H100 80GB HBM3: TRELLIS Image-to-3D 167.2 assets/hour, TRELLIS.2 Image-to-3D 87.9 (4B), TripoSG 559.4 (1B), TripoSR 2107.4. 2 of 3 beat TRELLIS Image-to-3D here. These aren't like-for-like: TripoSR returns a rough vertex-coloured mesh and TripoSG bare geometry, while both TRELLIS numbers are timed to a finished, textured GLB.
Cost on a rented GPU. 1,000 assets of TRELLIS Image-to-3D: $5.87 on a L40S ($0.79/hr, 7.4 hours), $12.78 on a H100 80GB HBM3 ($2.14/hr, 6.0 hours, 2.2x the cost).
TRELLIS Image-to-3D: cost per 1,000 assets on rented GPUs
| GPU | Cheapest rate | Speed (assets/hour) | Cost per 1,000 assets |
|---|---|---|---|
| NVIDIA L40S | $0.79/hr | 134.5 | $5.87 |
| NVIDIA L4 | $0.44/hr | 68.8 | $6.40 |
| NVIDIA A100 80GB SXM4 | $0.95/hr | 116.4 | $8.14 |
| NVIDIA H100 80GB HBM3 | $2.14/hr | 167.2 | $12.78 |
| NVIDIA H200 | $3.59/hr | 134.6 | $26.67 |
Cheapest hourly rate we track on RunPod and Vast.ai, divided by the measured speed. Startup time and storage are extra.
Speed tiers for TRELLIS Image-to-3D. 60-360 assets/hour: 6 (H100 80GB HBM3, H200, L40S). 360 assets/hour is ten seconds per mesh.
VRAM for TRELLIS Image-to-3D. Measured peak 11.7GB, so 16GB is the smallest common card size; smallest card it ran on: A10G (24GB).
TRELLIS v1 is still the fastest way to a textured GLB we measured: 167 an hour on an H100, 135 on an L40S, about twice TRELLIS.2's pace on the same card. It used about 12GB at its peak. Use it for background props and for drafts you'll pick from; use TRELLIS.2 for the assets people look at closely.
Warm worker: one pass over five game-asset pictures (rifle, pistol, magazine, crate, tree, as transparent cut-outs) to warm up, then the same five timed one after another. Each asset is timed from picture to a finished GLB: TRELLIS-image-large's default sampler settings, then the official to_glb export at 95% simplification with 1024px textures. Peak VRAM is the most PyTorch allocated during the timed pass. Datacenter cards on Modal, October 2026.