Kimi K3 (Max) — benchmark results

Provider: Moonshot. Released 2026-07-17. Access: API.

Unified ELO 2034 ± 52, rank #16 of 1776 rated models, from 11 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLM2014 Logic 2026-0774.8Median Score97.5
VoxelBench2040Rating95.7
OTIS Mock AIME 2024-2597.22Accuracy (%)93.6
SEAL - MCP Atlas82.3Score88.9
Epoch AI - ECI155.62ECI Score88.6
Epoch AI - Critpt23.4Score85.9
SimpleBench60.7Score (AVG@5)77.3
Chess Puzzles (Epoch AI)39Accuracy (%)75
FrontierMath - Tiers 1-3 (v2)72.18Accuracy (%, 285 private v2 problems)73
FrontierMath - Tier 4 (v2)39.02Accuracy (%, 41 private v2 problems)69.2
SimpleQA Verified42.7Accuracy (%)57.8

Interactive version: theaggregate.ai/model?slug=kimi-k3-max · How the rankings work · Data refreshed daily, snapshot 2026-07-22.