Claude Sonnet 4.5 (High) — benchmark results
Claude Sonnet 4.5 evaluated at the high reasoning-effort setting. Provider: Anthropic. Released 2025-09-29. Access: API.
Unified ELO 1973 ± 34, rank #35 of 1776 rated models, from 27 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| CLEM Clemscore | 90.1 | Clemscore (%) | 100 |
| CLEM Hot Air Balloon | 95.53 | Game Clemscore (%) | 100 |
| CLEM Wordle with Clue | 82.5 | Game Clemscore (%) | 100 |
| CLEM Wordle with Critic | 86.11 | Game Clemscore (%) | 100 |
| HAL GAIA Level 2 | 74.42 | Accuracy (%) | 100 |
| HAL SWE-bench Verified Mini | 72 | Score (%) | 100 |
| CLEM TextMapWorld | 87.53 | Game Clemscore (%) | 96.7 |
| CLEM Codenames | 73.85 | Game Clemscore (%) | 96.6 |
| CLEM AdventureGame | 97.5 | Game Clemscore (%) | 95 |
| CLEM MatchIt ASCII | 100 | Game Clemscore (%) | 94.8 |
| HAL GAIA | 70.91 | Accuracy (%) | 93.8 |
| HAL GAIA Level 1 | 77.36 | Accuracy (%) | 93.8 |
Interactive version: theaggregate.ai/model?slug=claude-sonnet-4-5-high · How the rankings work · Data refreshed daily, snapshot 2026-07-22.