GLM-5.2 (Max) — benchmark results
GLM-5.2 evaluated at the max reasoning-effort setting. Provider: Zhipu. Released 2026-06-13. Access: Open.
Unified ELO 1925 ± 25, rank #52 of 1776 rated models, from 65 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA TAU-2 Bench | 99.12 | Accuracy (%) | 99.9 |
| WebDev Arena | 1593.25 | Arena Score | 98.7 |
| AA Humanity's Last Exam | 40.13 | Accuracy (%) | 97.2 |
| Artificial Analysis Intelligence Index | 51.09 | Intelligence Index | 97 |
| Chatbot Arena (Code) | 1592 | Elo | 97 |
| AA CritPt | 20.86 | Accuracy (%) | 96.6 |
| AA Long Context Reasoning | 71.33 | Accuracy (%) | 95.4 |
| AA Terminal-Bench Hard | 50.76 | Accuracy (%) | 95 |
| AA SciCode | 50.46 | Accuracy (%) | 93.9 |
| AA GPQA Diamond | 89.49 | Accuracy (%) | 92.5 |
| LLM2014 Logic 2026-06 | 71.16 | Median Score | 92.5 |
| GDPval-AA | 1514 | Elo | 92.3 |
Interactive version: theaggregate.ai/model?slug=glm-5-2-max · How the rankings work · Data refreshed daily, snapshot 2026-07-22.