Qwen 2 VL 72B Instruct: benchmark results

Alibaba's open 72B vision-language flagship (September 2024) with dynamic-resolution image handling and 20-minute-plus video understanding. Provider: Alibaba. Released 2024-09-19. Access: Open.

Unified ELO 1551 ± 1, rank #430 of 1392 rated models, from 49 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLM Stats (EgoSchema)77.9Score (%)100
LLM Stats (TextVQA)85.5Score (%)100
Open LLM Leaderboard - MMLU-Pro52.41Score99.1
KOFFVQA - Object Attributes86.67Score (%)98.8
Open LLM Leaderboard - BBH56.31Score98.4
Open LLM Leaderboard - GPQA18.34Score96.8
Open LLM Leaderboard - MATH Level 534.44Score84.9
Open LLM Leaderboard - MuSR15.89Score84.2
KOFFVQA - Table Understanding84.33Score (%)81.5
LLM Stats (ChartQA)88.3Score (%)80
Enkrypt AI - Safety Risk21.62Risk Score76.7
LLM Stats (DocVQAtest)96.5Score (%)75

Interactive version: theaggregate.ai/model?slug=qwen-2-vl-72b-instruct · How It Works · Data refreshed daily, snapshot 2026-09-05.