Qwen 3.6 27B (Non-reasoning): benchmark results
Qwen 3.6 27B evaluated with reasoning disabled. Provider: Alibaba. Released 2026-04-22. Access: Open.
Unified ELO 1594 ± 1, rank #458 of 1761 rated models, from 46 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| CritPt | 90 | Accuracy (self-reported) | 97.7 |
| CPTU Bench | 4.29 | Average Score (1-5) | 93.6 |
| AA TAU-2 Bench | 93.57 | Accuracy (%) | 90.5 |
| UGI - Natural Intelligence | 30.67 | NatInt Score | 75.3 |
| AA Omniscience - Software Engineering (SWE) - Rust | 56 | Accuracy (%) | 74.1 |
| AA GPQA Diamond | 82.93 | Accuracy (%) | 72.4 |
| Artificial Analysis Intelligence Index | 23.31 | Intelligence Index | 70.1 |
| UGI - Writing | 38.83 | Writing Score | 68.8 |
| AA Humanity's Last Exam | 15.06 | Accuracy (%) | 66.4 |
| AA Omniscience - Software Engineering (SWE) - TypeScript | 25.56 | Accuracy (%) | 65.9 |
| AA Omniscience - Software Engineering (SWE) - PHP | 26 | Accuracy (%) | 65 |
| AA CritPt | 0.86 | Accuracy (%) | 62.5 |
Interactive version: theaggregate.ai/model?slug=qwen-3-6-27b-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-05.