Open LMM Reasoning - MathVision - arithmetic — leaderboard

Metric: Accuracy (%). Source: huggingface.co. 100 models tracked.

Top models

#ModelScore
1GPT-4.1 (2025-04-14)66.4
2Doubao-1.5-Pro65.7
3GPT-4.1 Mini62.9
4GPT-4o ChatGPT60
5Gemini 2.0 Flash57.1
6InternVL3-78B55.7
7Claude 3.7 Sonnet55.7
8Claude 3.5 Sonnet (20241022)53.6
9Gemini 1.5 Pro (002)53.6
10InternVL2.5-78B51.4
11InternVL3-14B50
12InternVL3-38B49.3
13Gemma 3 27B46.4
14QVQ-72B-Preview45
15GPT-4.1 Nano42.9

Interactive version: theaggregate.ai/benchmark?slug=open-lmm-reasoning-mathvision-arithmetic · How the rankings work · Data refreshed daily, snapshot 2026-07-22.