Claude Sonnet 5 (xHigh) — benchmark results
Claude Sonnet 5 evaluated at the xhigh reasoning-effort setting. Provider: Anthropic. Released 2026-06-30. Access: API.
Unified ELO 1908 ± 39, rank #57 of 1776 rated models, from 8 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| OTIS Mock AIME 2024-25 | 94.72 | Accuracy (%) | 87.2 |
| Epoch AI - ECI | 153.2 | ECI Score | 80 |
| VoxelBench | 1619 | Rating | 78.7 |
| Chess Puzzles (Epoch AI) | 35 | Accuracy (%) | 67.9 |
| LLM2014 Logic 2026-07 | 43.14 | Median Score | 62.5 |
| CursorBench 3.1 | 58.7 | Score (%) | 54.1 |
| Epoch AI - Cursorbench | 58.4 | Score | 50 |
| SimpleQA Verified | 25 | Accuracy (%) | 20.3 |
Interactive version: theaggregate.ai/model?slug=claude-sonnet-5-xhigh · How the rankings work · Data refreshed daily, snapshot 2026-07-22.