Grok 4.1 Fast: benchmark results

xAI's speed-oriented Grok 4.1 variant for agentic tool-calling workloads with a 2M-token context. Provider: xAI. Released 2025-11-19. Access: API.

Unified ELO 1626 ± 1, rank #190 of 1392 rated models, from 174 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Hack-Verifiable TextArena28.5Avg HR (self-reported)100
Ko-AgentBench - L5 Error Handling & Robustness34.75Adaptive Routing Score (%)100
ALL Bench LLM81.77Average Numeric Benchmark Score (%)97.4
MCP-Universe (LLM w/ Function Calls)60.58Avg Evaluator Score95
RAI-Bench - RAG Robustness (LC Abstention)95Rate (%)94.9
FormationEval97.6Accuracy (%)94.4
MLX Benchmark V2 - Coding63.64Accuracy (%)90
Pencil Puzzle Bench - Hitori26.7Direct-ask Success Rate (%)89
Pencil Puzzle Bench - Nurimaze6.7Direct-ask Success Rate (%)89
GSMA Open-Telco - TeleLogs71Score (%)88.6
Chess Bench LLM1400Lichess Rating87.7
RAI-Bench - RAG Robustness (LC Factuality)47Rate (%)87

Interactive version: theaggregate.ai/model?slug=grok-4-1-fast · How It Works · Data refreshed daily, snapshot 2026-09-05.