KataGo-Bench-1K — leaderboard

Go next-move prediction benchmark with 1,000 19x19 board positions sampled from KataGo annotation data. Models read the move history and rendered board state, then output the next move; scoring is accuracy against KataGo candidate moves.

Metric: Next-move Accuracy (%). Source: arxiv.org. Status: saturation imminent. 11 models tracked.

Top models

#ModelScore
1Claude 3.7 Sonnet34.3
2O1 Mini27.3
3DeepSeek R117.6
4Qwen 2.5 7B Instruct8
5Qwen 2.5 32B Instruct6.8
6DeepSeek R1 Distill Qwen 32B4.7
7Qwen 2.5 32B1.5
8Qwen 2.5 7B1.4
9DeepSeek-R1-Distill-Qwen-7B0.6

Interactive version: theaggregate.ai/benchmark?slug=katago-bench-1k · How the rankings work · Data refreshed daily, snapshot 2026-07-22.