Grok 4.1 Fast — benchmark results
xAI's speed-oriented Grok 4.1 variant for agentic tool-calling workloads with a 2M-token context. Provider: xAI. Released 2025-11-19. Access: API.
Unified ELO 1738 ± 18, rank #214 of 1776 rated models, from 150 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Hack-Verifiable TextArena | 28.5 | Avg HR (self-reported) | 100 |
| Ko-AgentBench - L5 Error Handling & Robustness | 34.75 | Adaptive Routing Score (%) | 100 |
| SpeechMap Compliance | 97.9 | % Requests Completed | 98.2 |
| ALL Bench LLM | 81.77 | Average Numeric Benchmark Score (%) | 97.4 |
| MCP-Universe (LLM w/ Function Calls) | 60.58 | Avg Evaluator Score | 95 |
| FormationEval | 97.6 | Accuracy (%) | 94.4 |
| AA-LCR | 68 | Score (self-reported) | 91.2 |
| MLX Benchmark V2 - Coding | 63.64 | Accuracy (%) | 90 |
| GSMA Open-Telco - TeleLogs | 71 | Score (%) | 89.5 |
| Pencil Puzzle Bench - Hitori | 26.7 | Direct-ask Success Rate (%) | 89 |
| Pencil Puzzle Bench - Nurimaze | 6.7 | Direct-ask Success Rate (%) | 89 |
| Chess Bench LLM | 1405 | Lichess Rating | 88.9 |
Interactive version: theaggregate.ai/model?slug=grok-4-1-fast · How the rankings work · Data refreshed daily, snapshot 2026-07-22.