Reasoning
Mathematical and scientific problem solving.
Model profile / Inception
Mercury 2.5. 260K context is advertised; the exact token convention is not stated. API model reference documents 65,536 maximum output tokens. Standard $0.20 input / $0.75 output per million tokens; temporary launch discount is separate. Provider 1,107 tokens/second is not a comparable AA 10K measurement.
Four areas of evidence
Mathematical and scientific problem solving.
Software engineering and programming.
Factual reliability and instruction following.
Tool use, workflows and professional tasks.
Current index snapshot · 2026-09-16. Each category uses its own scale; scores across categories are not directly comparable. Counts describe published catalogue evidence. A point lead does not establish superiority.
Performance in context
Mercury 2.5 has no current Capabilities rating. Explore the rated models below.
Lumina Capabilities Index
The current Capabilities cohort, with available operational measurements.
Chart loads as you explore
Lumina Capabilities Index
The current Capabilities cohort, with available operational measurements.
Chart loads as you explore
Explore the detail
No current index estimate. Available benchmark measurements and source records are retained below.
| Benchmark | Result | Published configuration | Source |
|---|---|---|---|
| No matching measurements. Try another search. | |||
No specialist task observations retained for this model.