Skip to content

The Aggregate

Unified LLM rankings, daily

Models What's New Trends Benchmarks Skill Maps
How It Works Guesswork Metrics
The Aggregate
Loading data...

OpenSkillEval — leaderboard

Metric: Overall avg. (self-reported). Source: benchmarklist.com. 10 models tracked.

Top models

#ModelScore
1Claude Opus 4.64.51
2GPT-5.54.47
3Claude Sonnet 4.64.43
4GLM-5.14.42
5GPT-5.24.03
6MiniMax-M2.74.02
7Gemini 3.1 Pro (Preview)4
8GPT-5.3 Codex3.76

Interactive version: theaggregate.ai/benchmark?slug=openskilleval · How the rankings work · Data refreshed daily, snapshot 2026-07-22.

Built by Mikhail Doroshenko — AI researcher, co-author of Humanity’s Last Exam. Independent project; no affiliation with any AI lab.