GLM-4.6V — benchmark results
Zhipu's open vision-language MoE (106B total/12B active) with native tool calling and 128K context (December 2025). Provider: Zhipu. Released 2025-12-09. Access: Open.
Unified ELO 1593 ± 37, rank #494 of 1776 rated models, from 48 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| MechVQA | 78.91 | Total (self-reported) | 93.8 |
| YapBench | 443.8 | YapIndex (lower is better) | 83.5 |
| LithoBench | 94.5 | MCQ$_{\mathrm{all}}$ (self-reported) | 83.3 |
| Ko-AgentBench - L5 Error Handling & Robustness | 26.01 | Adaptive Routing Score (%) | 78.6 |
| MathVision | 63.5 | Overall Accuracy (%) | 76.4 |
| AI for Education Visual Maths - Number and Operations | 59.46 | Accuracy (%) | 75.8 |
| AI for Education Visual Maths - Geometry | 54.62 | Accuracy (%) | 66.7 |
| AI for Education Visual Maths | 59.92 | Accuracy (%) | 65 |
| AI for Education Visual Reasoning - odd one out | 53.2 | Accuracy (%) | 62.9 |
| AI for Education Visual Maths - Algebra | 76.92 | Accuracy (%) | 60.8 |
| Chatbot Arena (Text) | 1377 | Elo | 58.8 |
| AI for Education Visual Maths - Measurement | 75.68 | Accuracy (%) | 58.3 |
Interactive version: theaggregate.ai/model?slug=glm-4-6v · How the rankings work · Data refreshed daily, snapshot 2026-07-22.