Qwen 2.5 Turbo — benchmark results

Alibaba's fast, low-cost proprietary Qwen2.5 API tier, whose Turbo update (November 2024) extended context to 1M tokens via sparse attention. Provider: Alibaba. Released 2024-09-19. Access: API.

Unified ELO 1407 ± 29, rank #1193 of 1776 rated models, from 11 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
AA MATH-50080.53Accuracy (%)47.3
AI Chess Leaderboard (Reasoning)574Elo30.9
AA MMLU-Pro63.31Accuracy (%)24.7
AI Chess Leaderboard (Continuation)438Elo23.5
BenchTable25.1Total Score (%)20.3
Artificial Analysis Intelligence Index6.3Intelligence Index18.3
AA GPQA Diamond41.01Accuracy (%)18
AA LiveCodeBench16.3Pass@1 (%)17
AA Humanity's Last Exam4.21Accuracy (%)15.6
AA SciCode15.28Accuracy (%)14.9
ResearchCodeBench8Task Success Rate (%)3.2

Interactive version: theaggregate.ai/model?slug=qwen-2-5-turbo · How the rankings work · Data refreshed daily, snapshot 2026-07-22.