OpenAI · stable
GPT-5.2 is a API model tracked via Artificial Analysis independent evaluations for capability comparison across coding, agents, and reasoning.
Specialist evidence
Source-native operational evidence for GPT-5.2. Exact configurations remain separate. Costs are comparable only within the same selected benchmark/evaluation, and these rows never enter Overall Score.
| Evaluation | Exact configuration | Performance | Cost / task | Tokens / task | Execution |
|---|---|---|---|---|---|
| SWE-bench Owner Leaderboards · owner-current · multilingual SWE-bench Owner Leaderboards · checked 2026-08-29 | gpt-5-2-livebench-2026-06-25-high GPT 5.2 (high) | 66.7% | $0.536 | — | 40.3 calls |