AI Chess Leaderboard (Reasoning) — leaderboard

LLMs play full chess games from the starting position. Ranked by Elo rating derived from head-to-head matches. Tests legal move generation, multi-step planning, and game-state tracking over 264 evaluated models.

Metric: Elo. Source: dubesor.de. Status: saturated. 300 models tracked.

Top models

#ModelScore
1Gemini 3.1 Pro (Preview)1922
2Gemini 3 Pro (Preview)1850
3Qwen 3 Max (Thinking)1800
4GPT-5 Codex1777
5Gemini 3.5 Flash1772
6Gemini 3.6 Flash1763
7GPT-5.1 Codex1743
8GPT-5.51720
9Claude Fable 51670
10GPT-5.6 Sol1604
11Grok 41596
12O31585
13GPT-51572
14Hy31532
15GPT-5.11527

Interactive version: theaggregate.ai/benchmark?slug=ai-chess-leaderboard-reasoning · How the rankings work · Data refreshed daily, snapshot 2026-07-22.