Qwen 2 VL 72B Instruct — benchmark results

Alibaba's open 72B vision-language flagship (September 2024) with dynamic-resolution image handling and 20-minute-plus video understanding. Provider: Alibaba. Released 2024-09-19. Access: Open.

Unified ELO 1539 ± 14, rank #649 of 1776 rated models, from 43 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLM Stats (EgoSchema)77.9Score (%)100
LLM Stats (TextVQA)85.5Score (%)100
Open LLM Leaderboard - MMLU-Pro52.41Score99.1
KOFFVQA - Object Attributes86.67Score (%)98.8
Open LLM Leaderboard - BBH56.31Score98.4
Open LLM Leaderboard - GPQA18.34Score96.8
Open LLM Leaderboard - MATH Level 534.44Score84.9
Open LLM Leaderboard - MuSR15.89Score84.2
KOFFVQA - Table Understanding84.33Score (%)81.5
LLM Stats (ChartQA)88.3Score (%)78.3
LLM Stats (DocVQAtest)96.5Score (%)75
KOFFVQA - Commonsense Reasoning83.11Score (%)74.1

Interactive version: theaggregate.ai/model?slug=qwen-2-vl-72b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.