Qwen 3.7 Plus (Non-reasoning): benchmark results
Provider: Alibaba. Released 2026-06-02. Access: API.
Unified ELO 1649 ± 1, rank #352 of 3078 rated models, from 15 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| BrainBench (EEG) - BrainAgent - Foundational Analysis | 73.08 | EEG analysis score (0-100) | 75 |
| BrainBench (EEG) - CodeAct - Neurocognitive Assessment | 64.17 | EEG analysis score (0-100) | 75 |
| BrainBench (EEG) - BrainAgent - Sleep Assessment | 68.08 | EEG analysis score (0-100) | 66.7 |
| BrainBench (EEG) - CodeAct - Foundational Analysis | 68.33 | EEG analysis score (0-100) | 66.7 |
| Epoch AI - GPQA Diamond | 81.82 | Accuracy (%) | 64.1 |
| OTIS Mock AIME 2024-25 | 80 | Accuracy (%) | 63.1 |
| BrainBench (EEG) - BrainAgent - Neurocognitive Assessment | 69.86 | EEG analysis score (0-100) | 58.3 |
| BrainBench (EEG) - BrainAgent - Overall | 67.25 | Difficulty-weighted EEG analysis score (0-100) | 58.3 |
| BrainBench (EEG) - CodeAct - Overall | 61.55 | Difficulty-weighted EEG analysis score (0-100) | 58.3 |
| FrontierMath - Tiers 1-3 (v2) | 34.39 | Accuracy (%, 285 private v2 problems) | 53.1 |
| BrainBench (EEG) - BrainAgent - Physiological Integration | 57.98 | EEG analysis score (0-100) | 50 |
| Epoch AI - Mystery Game Puzzles | 17 | Score | 47.2 |
Interactive version: theaggregate.ai/model?slug=qwen-3-7-plus-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-19.