Qwen 3.5 Plus — benchmark results
Alibaba's flagship Qwen 3.5 Plus API tier with thinking and non-thinking modes. Provider: Alibaba. Released 2026-02-16. Access: API.
Unified ELO 1719 ± 27, rank #243 of 1776 rated models, from 49 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AI for Education Pedagogy - Science | 95.08 | Accuracy (%) | 100 |
| ClawProBench | 64.19 | Final Score (self-reported) | 96.4 |
| AI for Education Visual Maths - Number and Operations | 72.97 | Accuracy (%) | 95.8 |
| AI for Education Pedagogy - Maths | 90.48 | Accuracy (%) | 94.5 |
| AI for Education Visual Maths - Statistics and Probability | 57.14 | Accuracy (%) | 94.2 |
| AI for Education Pedagogy - Secondary | 88.36 | Accuracy (%) | 93.9 |
| AI for Education Pedagogy | 88.88 | Accuracy (%) | 93.6 |
| AI for Education Visual Maths | 80.17 | Accuracy (%) | 91.7 |
| AI for Education Visual Maths - Geometry | 77.69 | Accuracy (%) | 91.7 |
| AI for Education Pedagogy - Technology | 85.85 | Accuracy (%) | 90.5 |
| AI for Education Visual Reasoning - pattern completion (linear) | 83.8 | Accuracy (%) | 90.3 |
| AI for Education Pedagogy - Social studies | 86.36 | Accuracy (%) | 88.6 |
Interactive version: theaggregate.ai/model?slug=qwen-3-5-plus · How the rankings work · Data refreshed daily, snapshot 2026-07-22.