GLM-4.7 (Thinking) — benchmark results
Provider: Zhipu. Released 2025-12-22. Access: Open.
Unified ELO 1696 ± 36, rank #280 of 1776 rated models, from 8 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LLM2014 Logic 2025-12 | 65.49 | Median Score | 86 |
| BenchTable | 69.6 | Total Score (%) | 85.5 |
| LLM2014 Logic 2026-01 | 62.94 | Median Score | 84.1 |
| ProfBench | 46.6 | Overall Rubric Score (%) | 60.1 |
| LiveMedBench | 13.35 | Overall Score (%) | 59.5 |
| LisanBench | 0.01 | Mean Path Length / Current Maximum | 36.2 |
| SecCodeBench | 52.29 | Total Score | 28.6 |
| HalluHard | 0.73 | Turn-1 Hallucination Rate | 19.4 |
Interactive version: theaggregate.ai/model?slug=glm-4-7-thinking · How the rankings work · Data refreshed daily, snapshot 2026-07-22.