Qwen 3.5 Plus (Thinking) — benchmark results

Qwen 3.5 Plus evaluated with thinking enabled. Provider: Alibaba. Released 2026-02-16. Access: API.

Unified ELO 1769 ± 24, rank #169 of 1776 rated models, from 26 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
SecCodeBench67.08Total Score97.1
BenchTable72.7Total Score (%)89.5
Vals AI LegalBench85.1Accuracy (%)88
Vals AI MedQA95.21Accuracy (%)87.2
Vals AI LiveCodeBench85.33Accuracy (%)81.9
Vals AI CorpFin v265.31Accuracy (%)80.3
Vals AI MMLU-Pro87.18Accuracy (%)79.3
Vals AI GPQA87.37Accuracy (%)78.3
LLM2014 Logic 2026-0253.8Median Score75.6
LLM2014 Logic 2026-0351.81Median Score70.7
CLBench19.8Solving Rate (%)68.6
Vals AI Finance Agent54.48Accuracy (%)68

Interactive version: theaggregate.ai/model?slug=qwen-3-5-plus-thinking · How the rankings work · Data refreshed daily, snapshot 2026-07-22.