OpenAI · stable
OpenAI's documented frontier API model with configurable reasoning and a 1.05-million-token context window.
Specialist evidence
Source-native operational evidence for GPT-5.5. Exact configurations remain separate. Costs are comparable only within the same selected benchmark/evaluation, and these rows never enter Overall Score.
| Evaluation | Exact configuration | Performance | Cost / task | Tokens / task | Execution |
|---|---|---|---|---|---|
| Artificial Analysis Coding Agents · 1.4 · composite Artificial Analysis Coding Agents · checked 2026-08-29 | gpt-5-5-medium Cursor CLI - GPT-5.5 (medium) | 47.1 | $2.00 | 3955818 | 77.8 steps |
| Artificial Analysis Coding Agents · 1.4 · composite Artificial Analysis Coding Agents · checked 2026-08-29 | gpt-5-5-medium Codex - GPT-5.5 (medium) | 55.3 | $2.65 | 6885561 | 78.5 steps |
| CursorBench · 3.2 · main CursorBench · checked 2026-08-29 | gpt-5-5-low GPT-5.5 Low | 46.6% | $0.980 | 5168 | 20 steps |
| CursorBench · 3.2 · main CursorBench · checked 2026-08-29 | gpt-5-5-medium GPT-5.5 Medium | 53.8% | $1.51 | 8522 | 25 steps |
| Terminal-Bench 2.1 · 2.1 · verified Terminal-Bench 2.1 · checked 2026-08-29 | gpt-5-5-xhigh Codex; reasoning xhigh; Terminal-Bench 2.1 verified submission. | 83.1% | $23.14 | — | — |
| Terminal-Bench 2.1 · 2.1 · verified Terminal-Bench 2.1 · checked 2026-08-29 | gpt-5-5-xhigh Terminus 2; reasoning xhigh; Terminal-Bench 2.1 verified submission. | 78.0% | $5.55 | — | — |