Qwen Plus (2025-07-28): benchmark results
Provider: Alibaba. Released 2025-01-25. Access: API.
Unified ELO 1807 ± 40, rank #144 of 2656 rated models, from 22 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| ReLE - Law and Civil Service | 82.7 | Accuracy (%) | 80.8 |
| ReLE - Language and Instruction Following | 70.1 | Accuracy (%) | 78.1 |
| ReLE - Finance | 82.8 | Accuracy (%) | 77.4 |
| ReLE - Medicine - Physician Exams | 85.2 | Accuracy (%) | 76.3 |
| ReLE - Education - Gaokao | 55.5 | Accuracy (%) | 72.5 |
| Tinybird AI SQL Benchmark - Success Rate | 100 | Questions answered with a valid query within 3 retries (%) | 72.5 |
| ReLE - Law - Lawyer Qualification Exam | 75.3 | Accuracy (%) | 72 |
| Kagi LLM Benchmark | 63.3 | Accuracy (%) | 69.7 |
| ReLE - Overall | 67.7 | Accuracy (%) | 68.9 |
| ReLE - Medicine and Mental Health | 82 | Accuracy (%) | 67.2 |
| ReLE - Reasoning - BBH | 80.1 | Accuracy (%) | 58.1 |
| ReLE - Education | 50.8 | Accuracy (%) | 55.9 |
Interactive version: theaggregate.ai/model?slug=qwen-plus-2025-07-28 · How It Works · Data refreshed daily, snapshot 2026-09-19.