O1 (2024-12-17) (High) — benchmark results
O1 (2024-12-17) evaluated at the high reasoning-effort setting. Provider: OpenAI. Released 2024-12-17. Access: API.
Unified ELO 1719 ± 45, rank #245 of 1776 rated models, from 11 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| UGI - Natural Intelligence | 63.08 | NatInt Score | 95.9 |
| UGI - Writing | 60.11 | Writing Score | 93.6 |
| Arena-Hard v2 (GPT-4.1 Judge) | 58.7 | Win Rate (%) | 85.2 |
| MATH Level 5 | 94.71 | Accuracy (%) | 84.3 |
| Arena-Hard v2 | 61 | Win Rate (%) | 77.8 |
| Aider polyglot coding leaderboard | 61.7 | Pass rate (%) | 71.2 |
| Arena-Hard Creative Writing | 59.9 | Win Rate (%) | 66.7 |
| UGI Leaderboard | 37.24 | UGI Score | 58.6 |
| SimpleBench | 40.1 | Score (AVG@5) | 39.8 |
| UGI - Willingness (W/10) | 2.5 | W/10 Score | 18.3 |
| Epoch AI - Apex Agents | 1.1 | Score | 1 |
Interactive version: theaggregate.ai/model?slug=o1-2024-12-17-high · How the rankings work · Data refreshed daily, snapshot 2026-07-22.