Qwen 3.6 27B — benchmark results
Alibaba's dense open Qwen 3.6 27B model, tuned for agentic coding. Provider: Alibaba. Released 2026-04-22. Access: Open.
Unified ELO 1691 ± 12, rank #289 of 1776 rated models, from 125 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| CheXpercept | 92.2 | Stage 1 (End-to-End) (self-reported) | 100 |
| LLM Stats (EmbSpatialBench) | 84.6 | Score (%) | 100 |
| LLM Stats (RefSpatialBench) | 70 | Score (%) | 100 |
| LLM Stats (VideoMME w sub.) | 87.7 | Score (%) | 100 |
| RefSpatialBench | 70 | RefSpatialBench (self-reported) | 100 |
| The Age of Curiosity Meets the Age of AI: Benc | 4.98 | Total (self-reported) | 100 |
| Done, But Not Sure | 37.8 | B All (self-reported) | 94.7 |
| LLM Stats (MVBench) | 75.5 | Score (%) | 93.8 |
| AI for Education Pedagogy - Primary | 92.96 | Accuracy (%) | 93.6 |
| AI for Education Pedagogy - Science | 92.35 | Accuracy (%) | 93.4 |
| AI for Education Pedagogy - Social studies | 87.27 | Accuracy (%) | 93.2 |
| AA-LCR | 68.7 | Score (self-reported) | 92 |
Interactive version: theaggregate.ai/model?slug=qwen-3-6-27b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.