Qwen 3.8 27B (Non-reasoning): benchmark results

Provider: Alibaba. Released 2026-08-14. Access: Open.

Unified ELO 1612 ± 1, rank #375 of 1761 rated models, from 23 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
BRIDGE Medical Leaderboard - Few-Shot55.67Average Performance (%)100
BRIDGE Medical Leaderboard47.74Average Performance (%)98.1
BRIDGE Medical Leaderboard - Zero-Shot44.6Average Performance (%)97.2
BRIDGE Medical Leaderboard - CoT42.94Average Performance (%)96.3
Artificial Analysis Intelligence Index26.45Intelligence Index75.2
AA Omniscience-7.95Score75
AA GPQA Diamond81.82Accuracy (%)69.9
AA Long Context Reasoning69.33Accuracy (%)65.6
AA Humanity's Last Exam12.14Accuracy (%)61.2
AA GDPval1143.29ELO59.5
AA CritPt0.29Accuracy (%)53.2
Tau3 Banking20Success Rate (%)52.5

Interactive version: theaggregate.ai/model?slug=qwen-3-8-27b-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-05.