GLM-5V Turbo (Reasoning): benchmark results
Provider: Zhipu. Released 2026-04-01. Access: API.
Unified ELO 1641 ± 1, rank #298 of 1919 rated models, from 20 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA TAU-2 Bench | 98.54 | Accuracy (%) | 99 |
| CritPt | 60 | Accuracy (self-reported) | 91.9 |
| Artificial Analysis Intelligence Index | 23.5 | Intelligence Index | 78.1 |
| AA Terminal-Bench Hard | 32.58 | Accuracy (%) | 76.6 |
| AA Omniscience - Health | 31 | Accuracy (%) | 75.6 |
| SpeechMap Compliance | 78.1 | % Requests Completed | 74.8 |
| AA Omniscience - Humanities & Social Sciences | 29.76 | Accuracy (%) | 74.5 |
| AA Omniscience - Science, Engineering & Mathematics | 36.94 | Accuracy (%) | 73.5 |
| AA IFBench | 61.09 | Accuracy (%) | 72.4 |
| AA-Omniscience Accuracy | 29.33 | Accuracy (%) | 72.2 |
| AA Omniscience - Software Engineering (SWE) | 40.34 | Accuracy (%) | 71.8 |
| AA GPQA Diamond | 80.91 | Accuracy (%) | 68.4 |
Interactive version: theaggregate.ai/model?slug=glm-5v-turbo-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-08.