Qwen Max — benchmark results

Alibaba's flagship Qwen Max model, represented here as a 700B-parameter closed model. Provider: Alibaba. Released 2025-04-28. Access: API.

Unified ELO 1465 ± 91, rank #932 of 1776 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
DuckDB-NSQL64Execution Accuracy (%)77.6
SpeechMap Compliance66.8% Requests Completed56.1
HSCodeComp - 10-digit3.8Exact Match Accuracy (%)46.2
HSCodeComp - 8-digit11.23Exact Match Accuracy (%)46.2
LLM Chess (Saplin)-98.9ELO42.9
SnakeBench20.3TrueSkill Rating41.3
HSCodeComp - 2-digit71.52Exact Match Accuracy (%)38.5
HSCodeComp - 4-digit48.58Exact Match Accuracy (%)30.8
HSCodeComp - 6-digit24.21Exact Match Accuracy (%)30.8
LiveCodeBench Pro274Rating (CF-style)1.9
PHYBench15.33EED Score0
SWE-Arena996Elo Score0

Interactive version: theaggregate.ai/model?slug=qwen-max · How the rankings work · Data refreshed daily, snapshot 2026-07-22.