Kimi K2 0905 — benchmark results
September 5, 2025 update of Moonshot's open Kimi K2 MoE with improved agentic coding and a 256K context. Provider: Moonshot. Released 2025-09-05. Access: Open.
Unified ELO 1685 ± 17, rank #301 of 1776 rated models, from 201 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AGC-Bench - future_ideas | 2.07 | Dataset z-score | 100 |
| AGC-Bench - hypogen | 2.19 | Dataset z-score | 100 |
| AGC-Bench - pron_vs_prompt | 1.41 | Dataset z-score | 100 |
| AGC-Bench - thenextchapter | 2.77 | Dataset z-score | 100 |
| Konkur 1404 - Art | 57.47 | Accuracy (%, text-only) | 100 |
| AGC-Bench - conceptual_design | 2.02 | Dataset z-score | 98.8 |
| AGC-Bench - cpers | 1.4 | Dataset z-score | 98.8 |
| AGC-Bench - cue_word_story | 1.16 | Dataset z-score | 98.8 |
| AGC-Bench - tinystories | 1.67 | Dataset z-score | 98.8 |
| AGC-Bench - rpgbench | 1.54 | Dataset z-score | 97.5 |
| AGC-Bench - tinyfabulist | 1.45 | Dataset z-score | 95.1 |
| Konkur 1404 - Mathematics | 52.17 | Accuracy (%, text-only) | 94.7 |
Interactive version: theaggregate.ai/model?slug=kimi-k2-0905 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.