Navigate the site or open source-checked model and benchmark registry records.
Navigate
Lumina Bench
News
Research
Independent rankings. Commercial relationships never affect scoring.
xAI · stable
xAI's documented general chat model with configurable reasoning and a one-million-token context window.
#36 · 65.1 score · Confidence B