Navigate the site or open source-checked model and benchmark registry records.
Navigate
Lumina Bench
News
Research
Independent rankings. Commercial relationships never affect scoring.
OpenAI · stable
o4-mini is tracked via Artificial Analysis independent evaluations for capability comparison across coding, agents, reasoning and related workloads.
#67 · 58.6 score · Confidence B