Grok 4.1 Fast — benchmark results

xAI's speed-oriented Grok 4.1 variant for agentic tool-calling workloads with a 2M-token context. Provider: xAI. Released 2025-11-19. Access: API.

Unified ELO 1738 ± 18, rank #214 of 1776 rated models, from 150 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Hack-Verifiable TextArena28.5Avg HR (self-reported)100
Ko-AgentBench - L5 Error Handling & Robustness34.75Adaptive Routing Score (%)100
SpeechMap Compliance97.9% Requests Completed98.2
ALL Bench LLM81.77Average Numeric Benchmark Score (%)97.4
MCP-Universe (LLM w/ Function Calls)60.58Avg Evaluator Score95
FormationEval97.6Accuracy (%)94.4
AA-LCR68Score (self-reported)91.2
MLX Benchmark V2 - Coding63.64Accuracy (%)90
GSMA Open-Telco - TeleLogs71Score (%)89.5
Pencil Puzzle Bench - Hitori26.7Direct-ask Success Rate (%)89
Pencil Puzzle Bench - Nurimaze6.7Direct-ask Success Rate (%)89
Chess Bench LLM1405Lichess Rating88.9

Interactive version: theaggregate.ai/model?slug=grok-4-1-fast · How the rankings work · Data refreshed daily, snapshot 2026-07-22.