Qwen 3.8 Flash (Max): benchmark results

Provider: Alibaba. Released 2026-08-26. Access: API.

Unified ELO 1716 ± 1, rank #97 of 3078 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
SuperCLUE General (July 2026) - Agentic Task Planning88.61Score82.4
SuperCLUE General (July 2026) - Hallucination Control86.08Score79.4
Bug Hunt Bench - VS Code Extension13Planted Bugs Fixed (out of 45)75.4
AI Coding Daily (OpenCode) - React-TS Code Quality17.67React-TS Code Quality (max 20) points, LLM-judged rubric sco71.1
Bug Hunt Bench26Planted Bugs Fixed (out of 105)69.4
AI Coding Daily (OpenCode) - Total47.22Total points (max 60)63.2
Bug Hunt Bench - LMS13Planted Bugs Fixed (out of 60)62.7
SuperCLUE General (July 2026) - Math Reasoning75.44Score61.8
SuperCLUE General (July 2026) - Overall66.41Score58.8
AI Coding Daily (OpenCode) - Laravel Code Quality16.35Laravel Code Quality (max 20) points, LLM-judged rubric scor31.6
SuperCLUE General (July 2026) - Precise Instruction Following30.48Score29.4
SuperCLUE General (July 2026) - Science Reasoning68.42Score23.5

Interactive version: theaggregate.ai/model?slug=qwen-3-8-flash-max · How It Works · Data refreshed daily, snapshot 2026-09-19.