GPT-4o (May '24) — benchmark results
Provider: OpenAI. Released 2024-05-13. Access: API.
Unified ELO 1492 ± 41, rank #872 of 1841 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA MATH-500 | 79.13 | Accuracy (%) | 45.6 |
| AA MMLU-Pro | 73.95 | Accuracy (%) | 44.5 |
| Epoch AI - Scicode | 30.9 | Score | 43.9 |
| AA LiveCodeBench | 33.44 | Pass@1 (%) | 39.9 |
| AA GPQA Diamond | 52.63 | Accuracy (%) | 29.5 |
| Artificial Analysis Intelligence Index | 8.6 | Intelligence Index | 28.1 |
| AA Humanity's Last Exam | 2.76 | Accuracy (%) | 0.5 |
Interactive version: theaggregate.ai/model?slug=gpt-4o-may-24 · How It Works · Data refreshed daily, snapshot 2026-07-25.