OpenAI · stable
GPT-5 mini is a API model tracked via Artificial Analysis independent evaluations for capability comparison across coding, agents, and reasoning.
Specialist evidence
Source-native operational evidence for GPT-5 mini. Exact configurations remain separate. Costs are comparable only within the same selected benchmark/evaluation, and these rows never enter Overall Score.
| Evaluation | Exact configuration | Performance | Cost / task | Tokens / task | Execution |
|---|---|---|---|---|---|
| SWE-bench Owner Leaderboards · owner-current · bash-only SWE-bench Owner Leaderboards · checked 2026-08-29 | gpt-5-mini-medium GPT 5 mini (2025-08-07) (medium) | 59.8% | $0.035 | — | 14.5 calls |
| SWE-bench Owner Leaderboards · owner-current · multilingual SWE-bench Owner Leaderboards · checked 2026-08-29 | gpt-5-mini-high GPT 5 mini | 39.7% | $0.052 | — | 30 calls |
| SWE-bench Owner Leaderboards · owner-current · verified SWE-bench Owner Leaderboards · checked 2026-08-29 | gpt-5-mini-medium mini-SWE-agent + GPT 5 mini (2025-08-07) (medium) | 59.8% | $0.035 | — | 14.5 calls |