Grok 4.20 (Non-reasoning) — benchmark results

Grok 4.20 evaluated with reasoning disabled. Provider: xAI. Released 2026-03-10. Access: API.

Unified ELO 1634 ± 43, rank #408 of 1776 rated models, from 13 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Design Arena (Website)1240Elo72.4
Design Arena (Data Viz)1234Elo71
BenchTable58.7Total Score (%)68.6
Design Arena (UI Components)1227Elo66.2
Design Arena (Game Dev)1226Elo64.3
Design Arena (3D)1207Elo62.9
AI Chess Leaderboard (Reasoning)664Elo51.2
CLBench17.7Solving Rate (%)45.7
LLM Chess (Saplin)-125ELO35.7
Opper TaskBench82Avg Task Score (%)33.5
AI Chess Leaderboard (Continuation)471Elo29.6
Design Arena (SVG)1103Elo26.3

Interactive version: theaggregate.ai/model?slug=grok-4-20-non-reasoning · How the rankings work · Data refreshed daily, snapshot 2026-07-22.