Kimi Code Bench v2 leaderboard
2 ranked models · higher is better · labels show rank and score
View accessible chart data
| Model | Rank | Provider | Score |
|---|---|---|---|
| Kimi K3 | #1 | Moonshot AI | 72.9% |
| Kimi K2.7 Code | #2 | Moonshot AI | 62% |
Coding · Benchmark profile
Moonshot internal end-to-end coding tasks used in the Kimi K3 launch table; retained reference-only.
Data verified 27 Jul 2026 · Methodology 1.8.0
Benchmark score on Kimi Code Bench v2
Kimi K3 leads at 72.9%, followed by Kimi K2.7 Code (62%).
2 ranked models · higher is better · labels show rank and score
| Model | Rank | Provider | Score |
|---|---|---|---|
| Kimi K3 | #1 | Moonshot AI | 72.9% |
| Kimi K2.7 Code | #2 | Moonshot AI | 62% |
One best score per model · higher is better
| Rank | Model | Provider | License | Evidence use | Score |
|---|---|---|---|---|---|
| #1 | Kimi K3 kimi-k3 | Moonshot AI | closed | Reference only | 72.9% |
| #2 | Kimi K2.7 Code kimi-k2.7-code | Moonshot AI | open | Estimated reference | 62% |
About Kimi Code Bench v2
Moonshot internal end-to-end coding tasks used in the Kimi K3 launch table; retained reference-only. Results stay tied to the exact model variant and evaluation system. Multiple systems for the same model use the best published score on this page; overall Lumina scoring uses the median of ranking-eligible rows.
Open benchmark source ↗FAQ
Moonshot internal end-to-end coding tasks used in the Kimi K3 launch table; retained reference-only.
Kimi K3 by Moonshot AI currently leads with 72.9%.
2 models in the LuminaBench cohort have a qualifying score on this benchmark.
No. This benchmark is display-only and does not enter the overall Lumina composite.
Related