Leaderboard ANIMALS
A text prompt turned directly into 3D.
Ranked by human votes · 340 cast · see the AI-judge board →
How ranking works
Rank groups models whose 95% bootstrap CIs overlap into the same rank (they are not statistically separable). BT = Bradley–Terry score (Elo-scaled); the bar shows the 95% CI with the point estimate marked. Every board ranks a SINGLE method — scores from different methods come from disconnected match pools and aren't comparable, so there is no cross-method ranking. Elo updates live per vote. Recompute in Admin.
Text→3D (native)
7 generators| Rank | Generator | BT score | 95% interval BT 264.2–1661.1 | Trend | Votes | Status |
|---|---|---|---|---|---|---|
| 1 | Tripo H3.1 text | 1268.6 | [863.6, 1661.1] | 5 | 25 more votes → firm ► | |
| 1 | Tripo P1 text | 1247.3 | [925.0, 1593.0] | 7 | 23 more votes → firm ► | |
| 1 | Rodin text via Replicate | 1136.9 | [662.9, 1628.0] | 6 | 24 more votes → firm ► | |
| 1 | Meshy v6 text | 1135.6 | [749.6, 1497.0] | 6 | 24 more votes → firm ► | |
| 1 | Hunyuan3D v3 text | 985.4 | [572.4, 1211.2] | 9 | 21 more votes → firm ► | |
| 1 | Hunyuan3D 3.1 text | 918.7 | [264.2, 1175.8] | 5 | 25 more votes → firm ► | |
| 1 | Rodin text via fal | 888.1 | [501.9, 1074.8] | 6 | 24 more votes → firm ► |
95% credible interval · point estimate · trend · firm a rank backed by 30+ votes; below that, Status counts the votes still needed · click a row for detail
Filters & bias audit
Bias audit — left(A) win rate 0.569 (≈0.50 = unbiased) · tie 0.059 · bad 0.191
VLM judge (Sonnet 4.6, multi-view) — automated LLM-judge rankings by paradigm
Loading automated rankings…