GPT-5.1 ChatGPT — benchmark results
Provider: OpenAI. Released 2025-11-12. Access: API.
Unified ELO 1703 ± 61, rank #265 of 1776 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AI Chess Leaderboard (Continuation) | 1363 | Elo | 93.4 |
| AI Chess Leaderboard (Reasoning) | 837 | Elo | 72.6 |
| Bullshit Benchmark | 36.4 | BS Detection Rate (%) | 71.9 |
| Chess Bench LLM | 545 | Lichess Rating | 56.6 |
| SvelteBench | 83.3 | Average pass@1 (%) | 36.6 |
| LLM Chess (Saplin) | -130.7 | ELO | 34.3 |
| SpeechMap Compliance | 40.8 | % Requests Completed | 18.2 |
Interactive version: theaggregate.ai/model?slug=gpt-5-1-chatgpt · How the rankings work · Data refreshed daily, snapshot 2026-07-22.