MEGA-Bench Task - Visual Prediction Rater Surface Normal Estimation — leaderboard

Metric: Task Score (%). Source: huggingface.co. 44 models tracked.

Top models

#ModelScore
1Gemini 2.5 Pro90.5
2Gemini 1.5 Pro (002)81
3InternVL3-38B73.8
4Gemini 2.0 Flash (Preview)69
5InternVL3-78B66.7
6Claude 3.5 Sonnet (20241022)66.7
7GPT-4o57.1
8Claude 3.5 Sonnet (20240620)57.1
9InternVL2.5-78B54.8
10Gemma 3 27B (IT)42.9
11Gemini 1.5 Flash (002)33.3
12Aquila-VL-2B26.2
13InternVL3-14B23.8
14Qwen 2 VL 72B16.7
15Pixtral-12B16.7

Interactive version: theaggregate.ai/benchmark?slug=mega-bench-task-visual-prediction-rater-surface-normal-estimation · How the rankings work · Data refreshed daily, snapshot 2026-07-22.