Grok 2 (Dec '24) — benchmark results
Provider: xAI. Released 2024-12-12. Access: API.
Unified ELO 1480 ± 36, rank #921 of 1841 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA MATH-500 | 77.8 | Accuracy (%) | 43.9 |
| AA MMLU-Pro | 70.88 | Accuracy (%) | 37.8 |
| Epoch AI - Scicode | 28.47 | Score | 37.6 |
| AA LiveCodeBench | 26.67 | Pass@1 (%) | 28.5 |
| AA GPQA Diamond | 51.01 | Accuracy (%) | 26.8 |
| Artificial Analysis Intelligence Index | 8.04 | Intelligence Index | 25 |
| AA Humanity's Last Exam | 3.8 | Accuracy (%) | 8.5 |
Interactive version: theaggregate.ai/model?slug=grok-2-dec-24 · How It Works · Data refreshed daily, snapshot 2026-07-25.