GLM-4.7 Flash (Non-reasoning): benchmark results
GLM-4.7 Flash evaluated with reasoning disabled. Provider: Zhipu. Released 2026-01-19. Access: Open.
Unified ELO 1484 ± 1, rank #942 of 1761 rated models, from 37 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA TAU-2 Bench | 91.81 | Accuracy (%) | 86.9 |
| UGI - Willingness (W/10) | 6.8 | W/10 Score | 67 |
| AA Omniscience - Software Engineering (SWE) - Julia | 12.5 | Accuracy (%) | 63.8 |
| AA IFBench | 46.26 | Accuracy (%) | 54.5 |
| AA Omniscience - Software Engineering (SWE) - Python | 18.09 | Accuracy (%) | 52.6 |
| AA Omniscience - Software Engineering (SWE) - Java | 15 | Accuracy (%) | 49.8 |
| UGI Leaderboard | 33.75 | UGI Score | 46.8 |
| Artificial Analysis Intelligence Index | 9.56 | Intelligence Index | 46.7 |
| UGI - Natural Intelligence | 20.96 | NatInt Score | 45.3 |
| AA Omniscience - Software Engineering (SWE) - PHP | 18 | Accuracy (%) | 43 |
| AA Omniscience - Software Engineering (SWE) - Kotlin | 14 | Accuracy (%) | 39.8 |
| AA Omniscience - Software Engineering (SWE) - C | 27 | Accuracy (%) | 38.2 |
Interactive version: theaggregate.ai/model?slug=glm-4-7-flash-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-05.