Qwen Plus: benchmark results

Provider: Alibaba. Released 2025-01-25. Access: API.

Unified ELO 1643 ± 20, rank #480 of 2656 rated models, from 48 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Software Engineering Arena - Model Arena1002Elo Rating86.4
Phare - Bias Resistance55.7Score (%)80.3
Tinybird AI SQL Benchmark - First-Attempt Success Rate98Questions answered with a valid query on the first attempt (78.1
Tinybird AI SQL Benchmark - Success Rate100Questions answered with a valid query within 3 retries (%)72.5
IKP - T299.5Accuracy (%)72.2
SnakeBench25.9TrueSkill Rating71
IKP - T484.57Accuracy (%)66.2
Phare - Average Safety68.85Score (%)63.6
VerdictBench60.55Accuracy (self-reported)62.5
MedAraBench61.8Overall Accuracy (%)60
IKP - T393.99Accuracy (%)56.1
Phare - Harm Resistance94.14Score (%)55.7

Interactive version: theaggregate.ai/model?slug=qwen-plus · How It Works · Data refreshed daily, snapshot 2026-09-19.