Qwen 3.6 Plus (High): benchmark results

Provider: Alibaba. Released 2026-04-02. Access: API.

Unified ELO 1675 ± 32, rank #430 of 2133 rated models, from 15 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
InfiniteBM Heads-Up No-Limit Hold'em1330.72Arena Elo (elo)81.8
LivingScreen44.9Success rate (%), mean of the three task tiers; LivingScreen50
LivingScreen - Closed-Loop Application22.8Success rate (%) on L3 fact-checking, content moderation and50
LivingScreen - GUI Action72.8Success rate (%) on L1 atomic GUI actions (liking, collectin50
BaFCo - Coarse Layout Analysis (CoT)6.93Mean average precision at IoU 0.3 (0-100): class-wise averag44.4
BaFCo - Coarse Layout Analysis (Zero-shot)5.89Mean average precision at IoU 0.3 (0-100): class-wise averag44.4
BaFCo - Layout Analysis (CoT)1.66Mean average precision at IoU 0.3 (0-100): class-wise averag44.4
BaFCo - Layout Analysis (Zero-shot)1.75Mean average precision at IoU 0.3 (0-100): class-wise averag44.4
BackendForge (Base Oracle)23.21Task success rate (%; base oracle of 7,250 items written bef42.3
The Aggregate - Long-tail Book Recall31.9Accuracy (%)29.4
LivingScreen - Cross-Source Understanding39.1Success rate (%) on L2 multiple-choice questions linking vid25
The Aggregate - Exotic Fact Recall31.8Accuracy (%)18.2

Interactive version: theaggregate.ai/model?slug=qwen-3-6-plus-high · How It Works · Data refreshed daily, snapshot 2026-10-11.