GLM-4.6V: benchmark results

Zhipu's open vision-language MoE (106B total/12B active) with native tool calling and 128K context (December 2025). Provider: Zhipu. Released 2025-12-09. Access: Open.

Unified ELO 1579 ± 1, rank #323 of 1392 rated models, from 77 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
MechVQA78.91Total (self-reported)93.8
LithoBench94.5MCQ$_{\mathrm{all}}$ (self-reported)83.3
Ko-AgentBench - L5 Error Handling & Robustness26.01Adaptive Routing Score (%)78.6
MathVision63.5Overall Accuracy (%)76.9
SnakeBench26.1TrueSkill Rating72.6
CULTURE-MT16.07Ineff. (%) (self-reported)71.4
AI for Education Visual Maths - Number and Operations59.46Accuracy (%)65.5
Chatbot Arena (Text - Chinese)1429Arena Score60.7
AI for Education Visual Reasoning - odd one out53.2Accuracy (%)60
Chatbot Arena (Text - Creative Writing)1349Arena Score59.1
LongWebBench5.96Overall (self-reported)58.3
AI for Education Visual Maths - Geometry54.62Accuracy (%)58.1

Interactive version: theaggregate.ai/model?slug=glm-4-6v · How It Works · Data refreshed daily, snapshot 2026-09-05.