← Back to project

Image-to-3D Pilot20 · End-to-End, Shape, and Texture

TRELLIS.2-4B vs Hunyuan3D-2.1 vs S3-T50 · 2026-08-13
image-to-3dbenchmarkToys4Kpilot20

TL;DR

S3-T50 matches Hunyuan on image-to-geometry alignment and beats it on both released X-Ray shape metrics.
TRELLIS.2 remains the clear geometry leader: CD 0.142 and F-score 0.770.
S3 leads all five texture-proxy metrics, but the GT latent reconstruction shares its SC-VAE decoder; this is diagnostic, not a paper-equivalent texture claim.

1 · One frozen set, three independent scorecards

Three released evaluator families keep semantic alignment, GT geometry fidelity, and GT-geometry texture fidelity separate.

2 · End-to-end semantic alignment · HY3D-Bench

ModelNULIP-2-I ↑Uni3D-I ↑
TRELLIS.2-4B200.15190.2546
Hunyuan3D-2.1200.16900.3555
S3-T50200.17080.3520
S3 edges Hunyuan on ULIP-2-I; Hunyuan edges S3 on Uni3D-I; TRELLIS.2 dominates both GT-shape metrics.

3 · Shape fidelity · X-Ray normalized evaluation

ModelNChamfer ↓F-score ↑
TRELLIS.2-4B200.14180.7696
Hunyuan3D-2.1200.34340.3623
S3-T50200.31810.4440

4 · Texture isolation · LumiTex metric stack

Proxy GT: decoded Toys4K shape/PBR latents. The original PBR .blend assets and LumiTex 133-case test set are unavailable.
ModelNFID ↓CLIP-FID ↓CMMD ↓CLIP-I ↑LPIPS ↓
TRELLIS.2-4B20417.3540.613.4870.84940.6640
Hunyuan3D-2.120420.1539.373.4610.85220.6640
S3-T5020385.0432.053.2340.87940.6184
S3 wins every metric against the decoded-latent proxy; the direction is consistent but the magnitude is not paper-comparable.
LumiTex's standard-mesh UV rebake exposes appearance differences on four representative cases.

5 · Runtime and frozen provenance

ModelMedian s/caseEstimated warm 20-case timeImplementation
TRELLIS.2-4B65.621.9 minmicrosoft/TRELLIS.2 @ 2ed033d
Hunyuan3D-2.1149.649.9 minTencent-Hunyuan/Hunyuan3D-2.1 @ 82920d
S3-T50106.835.6 minSS-11K / shape-17K / texture-17K
TRELLIS.2 is the fastest complete two-pass pipeline; Hunyuan's six-view PBR paint is the main runtime bottleneck.

6 · Promotion gates

Official upstream scorer functions · fixed seed 0 · 20 category-balanced Toys4K cases · H200 inference