O1 Pro — benchmark results
OpenAI's pro-tier o1 variant using more inference compute. Provider: OpenAI. Released 2024-12-05. Access: API.
Unified ELO 1608 ± 35, rank #463 of 1776 rated models, from 24 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| OpenAI ChatGPT Pro - AIME 2024 | 86 | pass@1 accuracy (%) | 100 |
| OpenAI ChatGPT Pro - AIME 2024 4/4 Reliability | 80 | 4/4 reliability (%) | 100 |
| OpenAI ChatGPT Pro - Codeforces | 90 | pass@1 percentile | 100 |
| OpenAI ChatGPT Pro - Codeforces 4/4 Reliability | 75 | 4/4 reliability percentile | 100 |
| OpenAI ChatGPT Pro - GPQA Diamond | 79 | pass@1 accuracy (%) | 100 |
| OpenAI ChatGPT Pro - GPQA Diamond 4/4 Reliability | 74 | 4/4 reliability (%) | 100 |
| SEAL - Agentic Tool Use (Enterprise) | 67.01 | Score | 93.9 |
| SEAL - Agentic Tool Use (Chat) | 61.42 | Score | 90.9 |
| TrackingAI IQ Test | 80.39 | IQ Test Score (%) | 73.3 |
| Visual-Language Understanding | 47.32 | Score (self-reported) | 71.7 |
| SEAL - VISTA | 47.32 | Score | 71 |
| LLM Stats (AIME 2024) | 86 | Score (%) | 69.8 |
Interactive version: theaggregate.ai/model?slug=o1-pro · How the rankings work · Data refreshed daily, snapshot 2026-07-22.