Qwen Plus (2025-07-28): benchmark results

Provider: Alibaba. Released 2025-01-25. Access: API.

Unified ELO 1807 ± 40, rank #144 of 2656 rated models, from 22 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
ReLE - Law and Civil Service82.7Accuracy (%)80.8
ReLE - Language and Instruction Following70.1Accuracy (%)78.1
ReLE - Finance82.8Accuracy (%)77.4
ReLE - Medicine - Physician Exams85.2Accuracy (%)76.3
ReLE - Education - Gaokao55.5Accuracy (%)72.5
Tinybird AI SQL Benchmark - Success Rate100Questions answered with a valid query within 3 retries (%)72.5
ReLE - Law - Lawyer Qualification Exam75.3Accuracy (%)72
Kagi LLM Benchmark63.3Accuracy (%)69.7
ReLE - Overall67.7Accuracy (%)68.9
ReLE - Medicine and Mental Health82Accuracy (%)67.2
ReLE - Reasoning - BBH80.1Accuracy (%)58.1
ReLE - Education50.8Accuracy (%)55.9

Interactive version: theaggregate.ai/model?slug=qwen-plus-2025-07-28 · How It Works · Data refreshed daily, snapshot 2026-09-19.