GRAB-Lite — leaderboard

GRaph Analysis Benchmark: 500 task-balanced questions testing multimodal models on graph/figure analysis — estimating means, intercepts, correlations, and transforms across synthetic and realistic figures.

Metric: Overall Score. Source: grab-benchmark.github.io. Status: saturation imminent. 38 models tracked.

Top models

#ModelScore
1Claude Fable 574
2GPT-5.571.8
3GPT-5.471
4Gemini 3.5 Flash63
5GPT-5.4 Mini62
6Claude Opus 4.860.6
7Gemini 3.1 Pro (Preview)59.4
8GPT-5.259
9Claude Opus 4.758.2
10Gemini 3 Flash58.2
11Gemini 3 Pro53.4
12Gemini 3.1 Flash Lite47.8
13Claude Sonnet 4.646.6
14GPT-546.2
15Claude Opus 4.645.6

Interactive version: theaggregate.ai/benchmark?slug=grab-lite · How the rankings work · Data refreshed daily, snapshot 2026-07-22.