Leaderboard PLANTS

A text prompt turned directly into 3D.

Ranked by human votes · 340 cast · see the AI-judge board →

Share Post on X
How ranking works

Rank groups models whose 95% bootstrap CIs overlap into the same rank (they are not statistically separable). BT = Bradley–Terry score (Elo-scaled); the bar shows the 95% CI with the point estimate marked. Every board ranks a SINGLE method — scores from different methods come from disconnected match pools and aren't comparable, so there is no cross-method ranking. Elo updates live per vote. Recompute in Admin.

Text→3D (native)

7 generators
Rank Generator BT score 95% interval BT 339.7–1554.9 Trend Votes Status
1 Meshy v6 text papercode 1294.9
[1185.9, 1554.9]
14 16 more votes → firm
1 Tripo H3.1 text papercode 1151.6
[942.5, 1346.6]
16 14 more votes → firm
1 Tripo P1 text papercode 1135.7
[889.5, 1394.7]
13 17 more votes → firm
1 Hunyuan3D 3.1 text papercode 1086.6
[886.2, 1238.9]
16 14 more votes → firm
1 Hunyuan3D v3 text papercode 1076.8
[881.0, 1258.6]
13 17 more votes → firm
2 Rodin text via Replicate papercode 898.5
[448.7, 1180.4]
6 24 more votes → firm
6 Rodin text via fal papercode 593.3
[339.7, 672.6]
10 20 more votes → firm

95% credible interval · point estimate · trend · firm a rank backed by 30+ votes; below that, Status counts the votes still needed · click a row for detail

Filters & bias audit

Bias audit — left(A) win rate 0.569 (≈0.50 = unbiased) · tie 0.059 · bad 0.191

VLM judge (Sonnet 4.6, multi-view) — automated LLM-judge rankings by paradigm

Loading automated rankings…