Kimi K2.6 (Non-reasoning): benchmark results
Kimi K2.6 evaluated with reasoning disabled. Provider: Moonshot. Released 2026-04-20. Access: Open.
Unified ELO 1613 ± 1, rank #373 of 1761 rated models, from 36 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA Omniscience - Software Engineering (SWE) - Rust | 74 | Accuracy (%) | 95.1 |
| UGI - Writing | 64.19 | Writing Score | 95 |
| UGI - Natural Intelligence | 59.46 | NatInt Score | 93.7 |
| AA Omniscience - Software Engineering (SWE) - Go | 40 | Accuracy (%) | 93.2 |
| AA Omniscience - Software Engineering (SWE) - PHP | 48 | Accuracy (%) | 92.3 |
| AA Omniscience - Software Engineering (SWE) - Kotlin | 40 | Accuracy (%) | 91.6 |
| AA TAU-2 Bench | 93.86 | Accuracy (%) | 91.1 |
| AA Omniscience - Software Engineering (SWE) - C | 60 | Accuracy (%) | 90.8 |
| AA Omniscience - Software Engineering (SWE) - Python | 37 | Accuracy (%) | 89.4 |
| AA Omniscience - Software Engineering (SWE) - Java | 29 | Accuracy (%) | 88.9 |
| AA Omniscience - Software Engineering (SWE) - TypeScript | 43.33 | Accuracy (%) | 88.1 |
| AA Omniscience - Software Engineering (SWE) - Swift | 52 | Accuracy (%) | 86.7 |
Interactive version: theaggregate.ai/model?slug=kimi-k2-6-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-05.