MEGA-Bench Task - SciBench Fundamental Without Solution — leaderboard

Metric: Task Score (%). Source: huggingface.co. 44 models tracked.

Top models

#ModelScore
1Gemini 2.5 Pro57.1
2Gemini 2.0 Flash (Preview)42.9
3GPT-4o34.7
4Gemini 1.5 Flash (002)34.7
5Gemini 1.5 Pro (002)30.6
6Gemma 3 27B (IT)28.6
7Claude 3.5 Sonnet (20240620)28.6
8InternVL3-78B26.5
9GPT-4o Mini26.5
10InternVL2.5-78B26.5
11Claude 3.5 Sonnet (20241022)24.5
12Llama 4 Scout Base24.5
13InternVL3-38B22.4
14Qwen 2 VL 72B18.4
15Gemma 3 12B (IT)16.3

Interactive version: theaggregate.ai/benchmark?slug=mega-bench-task-scibench-fundamental-without-solution · How the rankings work · Data refreshed daily, snapshot 2026-07-22.