Anthropic · stable
Anthropic's Claude Opus 4.7 model tracked via the BenchLM July 2026 leaderboard for capability comparison.
Specialist evidence
Source-native operational evidence for Claude Opus 4.7. Exact configurations remain separate. Costs are comparable only within the same selected benchmark/evaluation, and these rows never enter Overall Score.
| Evaluation | Exact configuration | Performance | Cost / task | Tokens / task | Execution |
|---|---|---|---|---|---|
| Artificial Analysis Coding Agents · 1.4 · composite Artificial Analysis Coding Agents · checked 2026-08-29 | claude-opus-4-7-max Claude Code - Opus 4.7 (max) | 51.6 | $5.92 | 16129723 | 108 steps |
| Terminal-Bench 2.1 · 2.1 · verified Terminal-Bench 2.1 · checked 2026-08-29 | claude-opus-4-7-max Opus 4.7; reasoning=max; agent=Claude Code | 68.9% | — | — | — |
| Terminal-Bench 2.1 · 2.1 · verified Terminal-Bench 2.1 · checked 2026-08-29 | claude-opus-4-7-max Opus 4.7; reasoning=max; agent=Terminus 2 | 66.1% | — | — | — |