GLM-4.5 — benchmark results
Zhipu's open GLM-4.5 MoE model for reasoning, coding, and agentic tasks (355B total, 32B active). Provider: Zhipu. Released 2025-07-28. Access: Open.
Unified ELO 1642 ± 14, rank #389 of 1776 rated models, from 112 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| YapBench | 1427 | YapIndex (lower is better) | 100 |
| ZeroEval MATH-500 | 98.2 | MATH-500 Score | 93.5 |
| AI for Education Pedagogy - Social studies | 87.27 | Accuracy (%) | 93.2 |
| SEAL - Fortress | 59.58 | Score | 90.9 |
| AI for Education Pedagogy - Technology | 85.85 | Accuracy (%) | 90.5 |
| LLM Stats (AIME 2024) | 91 | Score (%) | 84.9 |
| WebCoderBench - Visual Experience | 87.42 | Score (%) | 84.6 |
| AI for Education SEND | 81.19 | Accuracy (%) | 84.2 |
| LiveMedBench | 22.46 | Overall Score (%) | 83.8 |
| AI for Education Pedagogy - Primary | 90.61 | Accuracy (%) | 82 |
| AI for Education Pedagogy - Science | 90.16 | Accuracy (%) | 81.8 |
| MathArena - AIME 2025 | 93.33 | Accuracy (%) | 80.6 |
Interactive version: theaggregate.ai/model?slug=glm-4-5 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.