Grok Beta — benchmark results
xAI's grok-beta, the Grok-2-class 128K-context preview model that launched the public xAI API beta in November 2024. Provider: xAI. Released 2024-11-04. Access: API.
Unified ELO 1506 ± 47, rank #779 of 1776 rated models, from 33 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| SpeechMap Compliance | 93.7 | % Requests Completed | 95.9 |
| TextClass Benchmark | 1741.94 | Meta-Elo (self-reported) | 91.6 |
| ProLLM - Function Calling | 91.8 | Score (%) | 90.7 |
| VNTL Leaderboard | 71.27 | Accuracy (%) | 89.5 |
| EvalPlus (HumanEval+ & MBPP+) | 73 | Pass@1 avg (%) | 86.3 |
| ForecastBench | 64.5 | Overall Score (higher is better) | 64.9 |
| ProLLM - Summarization | 74.1 | Score (%) | 61.2 |
| AI for Education Pedagogy - Technology | 80.19 | Accuracy (%) | 60.7 |
| HumanEval+ | 80.5 | HumanEval+ pass@1 (self-reported) | 56.5 |
| ProLLM - StackEval | 90.4 | Score (%) | 51.2 |
| NYT Connections Original | 23.7 | Score (%) | 48.3 |
| ProLLM - Q&A Assistant | 93.1 | Score (%) | 44.1 |
Interactive version: theaggregate.ai/model?slug=grok-beta · How the rankings work · Data refreshed daily, snapshot 2026-07-22.