OpenAI · stable
o3-mini is a proprietary model variant recorded in the BenchLM public dataset.
Specialist evidence
Source-native operational evidence for o3-mini. Exact configurations remain separate. Costs are comparable only within the same selected benchmark/evaluation, and these rows never enter Overall Score.
| Evaluation | Exact configuration | Performance | Cost / task | Tokens / task | Execution |
|---|---|---|---|---|---|
| SWE-bench Owner Leaderboards · owner-current · lite SWE-bench Owner Leaderboards · checked 2026-08-29 | o3-mini-default Agentless Lite + O3 Mini (20250214) | 32.3% | — | — | — |
| SWE-bench Owner Leaderboards · owner-current · verified SWE-bench Owner Leaderboards · checked 2026-08-29 | o3-mini-default Agentless Lite + O3 Mini (20250214) | 42.4% | — | — | — |