Qwen2.5-Math-7B-CFT — benchmark results

Provider: Alibaba. Released 2025-01-30. Access: Open.

Unified ELO 1332 ± 14, rank #1467 of 1776 rated models, from 179 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard - MATH Level 555.74Score99.3
Open Arabic LLM - Arabic MMLU Civics (High School)47.13Accuracy (%)73.8
Open Japanese LLM - MR82.2Score (%)49.5
Open Arabic LLM - Madinah QA Arabic Language (Grammar)30.68Accuracy (%)41
Open LLM Leaderboard - GPQA4.81Score41
Open LLM Leaderboard - BBH24.59Score37.7
Open Japanese LLM - Wiki Coreference SET F11.61Score (%)35.7
Open LLM Leaderboard - MMLU-Pro21.61Score34.7
Open Japanese LLM - Jsem Exact Match71.97Score (%)31.6
Open LLM Leaderboard - MuSR6.62Score31.4
Open Japanese LLM - Jmmlu Exact Match47.59Score (%)30.4
Open Arabic LLM - Arabic MMLU Math (Primary School)42.05Accuracy (%)29.3

Interactive version: theaggregate.ai/model?slug=qwen2-5-math-7b-cft · How the rankings work · Data refreshed daily, snapshot 2026-07-22.