About VoxelBench Text-Prompt Leaderboard
Definition and scoring
- Organisation
- VoxelBench
- Category
- Knowledge
- Version
- 2025
- Direction
- higher is better
- Ranking use
- Reference
- Contamination risk
- Unknown
A live human-preference benchmark where language models turn text prompts into voxel structures and voters compare anonymous builds from the same prompt. Every genuine source row stays tied to its exact model label, configuration, benchmark version and evaluation system. The summary chart shows one best compatible score per canonical product; Score 2.0 admits only explicitly mapped, frozen protocols and keeps incompatible configurations separate.
Open benchmark source ↗