Claude Opus 4.7 (Adaptive Reasoning, Max Effort): benchmark results
Claude Opus 4.7 evaluated in adaptive-reasoning mode at max effort. Provider: Anthropic. Released 2026-04-16. Access: API.
Unified ELO 1667 ± 1, rank #193 of 1761 rated models, from 30 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| ITBench-AA | 46.7 | Average Precision at Full Recall (self-reported) | 95.5 |
| AA Terminal-Bench Hard | 51.52 | Accuracy (%) | 95.4 |
| AA Omniscience | 27.27 | Score | 95.3 |
| UGI - Natural Intelligence | 62.94 | NatInt Score | 95 |
| AA Omniscience - Science, Engineering & Mathematics | 51.77 | Accuracy (%) | 94.3 |
| AA Omniscience - Software Engineering (SWE) | 75.3 | Accuracy (%) | 93.8 |
| UGI - Writing | 61.4 | Writing Score | 93.6 |
| Artificial Analysis Intelligence Index | 44.29 | Intelligence Index | 93.5 |
| AA Humanity's Last Exam | 42.31 | Accuracy (%) | 93.3 |
| UGI Leaderboard | 51.79 | UGI Score | 92.9 |
| AA GPQA Diamond | 91.41 | Accuracy (%) | 92.5 |
| AA Omniscience - Business | 42.25 | Accuracy (%) | 92.2 |
Interactive version: theaggregate.ai/model?slug=claude-opus-4-7-adaptive-reasoning-max-effort · How It Works · Data refreshed daily, snapshot 2026-09-05.