GLM-4.7 — benchmark results
Zhipu GLM-4.7 model row. Provider: Zhipu. Released 2025-12-22. Access: Open.
Unified ELO 1720 ± 11, rank #241 of 1776 rated models, from 260 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| OpenEvals - EvasionBench | 82.91 | Accuracy (%) | 100 |
| AGC-Bench - historical_analogy | 1.94 | Dataset z-score | 99.4 |
| FormationEval | 98.6 | Accuracy (%) | 98.6 |
| AGC-Bench - c3_crosstalk | 1.42 | Dataset z-score | 97.6 |
| CLEM Hot Air Balloon | 90.89 | Game Clemscore (%) | 95.8 |
| AGC-Bench - arn | 0.86 | Dataset z-score | 95.7 |
| AGC-Bench - schnovel | 1.56 | Dataset z-score | 95.7 |
| OpenCompass Research - LiveCodeBench v6 | 83.8 | Score (%) | 95.7 |
| AGC-Bench - Problem Solving | 0.72 | JRT z-score | 95.1 |
| AGC-Bench - lcc_metaphor | 1.21 | Dataset z-score | 95.1 |
| AGC-Bench - simile_generation | 1.35 | Dataset z-score | 95 |
| AGC-Bench - fig_qa | 2.38 | Dataset z-score | 94.5 |
Interactive version: theaggregate.ai/model?slug=glm-4-7 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.