Skip to content

The Aggregate

Unified LLM rankings, daily

Models What's New Trends Benchmarks Skill Maps
How It Works Guesswork Metrics
The Aggregate
Loading data...

LLM Stats (Claw-Eval) — leaderboard

Metric: Score (%). Source: llm-stats.com. 13 models tracked.

Top models

#ModelScore
1Kimi K2.680.9
2GLM-5V Turbo75
3MiniMax-M374.5
4Hy368.5
5Qwen 3.7 Max65.2
6MiMo-V2.5-Pro64
7MiMo-V2.563.2
8Qwen 3.7 Plus62.7
9MiMo-V2-Pro61.5
10Qwen 3.6 27B60.6
11Qwen 3.6 Plus58.7
12MiMo-V2-Omni54.8
13Qwen 3.6 35B A3B50

Interactive version: theaggregate.ai/benchmark?slug=llm-stats-claw-eval · How the rankings work · Data refreshed daily, snapshot 2026-07-22.

Built by Mikhail Doroshenko — AI researcher, co-author of Humanity’s Last Exam. Independent project; no affiliation with any AI lab.