Anthropic · stable
Anthropic's documented model for complex agentic coding and enterprise work.
Specialist evidence
Source-native operational evidence for Claude Opus 4.8. Exact configurations remain separate. Costs are comparable only within the same selected benchmark/evaluation, and these rows never enter Overall Score.
| Evaluation | Exact configuration | Performance | Cost / task | Tokens / task | Execution |
|---|---|---|---|---|---|
| Artificial Analysis Coding Agents · 1.4 · composite Artificial Analysis Coding Agents · checked 2026-08-29 | claude-opus-4-8-epoch-claude-opus-4-8-low Claude Code - Opus 4.8 (low) | 49.0 | $2.18 | 5214050 | 68.2 steps |
| Artificial Analysis Coding Agents · 1.4 · composite Artificial Analysis Coding Agents · checked 2026-08-29 | claude-opus-4-8-high Claude Code - Opus 4.8 (high) | 57.6 | $3.78 | 9188691 | 104.8 steps |
| Artificial Analysis Coding Agents · 1.4 · composite Artificial Analysis Coding Agents · checked 2026-08-29 | claude-opus-4-8-max Claude Code - Opus 4.8 (max) | 62.1 | $7.72 | 17920579 | 165 steps |
| CursorBench · 3.2 · main CursorBench · checked 2026-08-29 | claude-opus-4-8-epoch-claude-opus-4-8-low Opus 4.8 Low | 53.1% | $2.02 | 19624 | 27 steps |
| CursorBench · 3.2 · main CursorBench · checked 2026-08-29 | claude-opus-4-8-high Opus 4.8 High | 58.0% | $3.15 | 33548 | 33 steps |
| CursorBench · 3.2 · main CursorBench · checked 2026-08-29 | claude-opus-4-8-max Opus 4.8 Max | 62.3% | $5.77 | 71411 | 44 steps |
| Terminal-Bench 2.1 · 2.1 · verified Terminal-Bench 2.1 · checked 2026-08-29 | claude-opus-4-8-high Claude Code; reasoning high; Terminal-Bench 2.1 verified submission. | 78.9% | $3.22 | — | — |