Mizan LLM Leaderboard — leaderboard

Metric: Average Score (0-100). Source: huggingface.co. 45 models tracked.

Top models

#ModelScore
1O374.7
2Gemini 2.5 Pro73.3
3GPT-5 Mini71.7
4Claude 3.7 Sonnet (20250219)71.3
5GPT-4.169.9
6Gemini 2.5 Flash69.9
7Gemini 2.0 Flash68.9
8GPT-4o68.8
9GPT-5 Nano66.4
10Gemini 2.0 Flash Lite66.4
11GPT-OSS-120B66.3
12GPT-4.1 Mini65.6
13GPT-5 Mini (Minimal)65.6
14DeepSeek Reasoner65.5
15DeepSeek V3 Chat64.6

Interactive version: theaggregate.ai/benchmark?slug=mizan-llm-leaderboard · How the rankings work · Data refreshed daily, snapshot 2026-07-22.