Qwen2.5-Math-7B-Instruct: benchmark results
Provider: Alibaba. Released 2024-09-19. Access: Open.
Unified ELO 1342 ± 1, rank #1281 of 1392 rated models, from 94 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MATH Level 5 | 58.08 | Score | 99.6 |
| Omni-MATH | 33.22 | Overall Accuracy (%) | 71.4 |
| MERA - SimpleAr | 98.9 | EM (%) | 63.7 |
| U-MATH - Sequences & Series | 51.95 | Accuracy (%) | 54.5 |
| MERA - MathLogicQA | 47.86 | Accuracy (%) | 53.6 |
| MERA - ruMultiAr | 32.32 | EM (%) | 53.1 |
| Open Japanese LLM - Jsick Exact Match | 78.71 | Score (%) | 51.7 |
| U-MATH - Precalculus | 76.25 | Accuracy (%) | 51.5 |
| MERA - RWSD | 55.38 | Accuracy (%) | 45.7 |
| U-MATH | 45.45 | Accuracy (%) | 45.5 |
| MERA - ruModAr | 50.85 | EM (%) | 43.1 |
| Open Japanese LLM - MR | 76.6 | Score (%) | 43 |
Interactive version: theaggregate.ai/model?slug=qwen2-5-math-7b-instruct · How It Works · Data refreshed daily, snapshot 2026-09-05.