Qwen3.5 Omni Flash — benchmark results
Alibaba's low-latency Flash tier of the omni-modal Qwen3.5-Omni family: native text/image/audio/video input with real-time speech output (March 2026). Provider: Alibaba. Released 2026-03-30. Access: API.
Unified ELO 1525 ± 53, rank #703 of 1776 rated models, from 36 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA TAU-2 Bench | 84.5 | Accuracy (%) | 75.7 |
| OmniGAIA - Overall | 33.9 | Overall Accuracy (%) | 73.3 |
| AA GPQA Diamond | 74.24 | Accuracy (%) | 60.4 |
| AA-LCR | 44 | Score (self-reported) | 60.3 |
| Artificial Analysis Intelligence Index | 18.99 | Intelligence Index | 57.2 |
| AA Long Context Reasoning | 44 | Accuracy (%) | 53.8 |
| AA Humanity's Last Exam | 7.09 | Accuracy (%) | 51 |
| AA Omniscience - Software Engineering (SWE) - Rust | 50 | Accuracy (%) | 50.1 |
| AA Omniscience - Software Engineering (SWE) - Dart | 20 | Accuracy (%) | 47.9 |
| AA MMMU-Pro | 64.74 | Accuracy (%) | 45.6 |
| AA Terminal-Bench Hard | 8.33 | Accuracy (%) | 41.8 |
| AA Omniscience - Law | 9.2 | Accuracy (%) | 36.4 |
Interactive version: theaggregate.ai/model?slug=qwen3-5-omni-flash · How the rankings work · Data refreshed daily, snapshot 2026-07-22.