Navigate the site or open source-checked model and benchmark registry records.
Navigate
Lumina Bench
News
Research
Independent rankings. Commercial relationships never affect scoring.
Anthropic · stable
Anthropic's Claude Sonnet 4.6 model tracked via the BenchLM July 2026 leaderboard for capability comparison.
#54 · 61.3 score · Confidence B