Claude Sonnet 5 (Max) — benchmark results

Claude Sonnet 5 evaluated at the max reasoning-effort setting. Provider: Anthropic. Released 2026-06-30. Access: API.

Unified ELO 2007 ± 34, rank #22 of 1776 rated models, from 9 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Vals AI ProofBench66Accuracy (%)87.1
Epoch AI - Scicode53.59Score80.9
Epoch AI - ECI153.2ECI Score80
Epoch AI - Critpt16.86Score77.5
CursorBench 3.161.5Score (%)64.9
Epoch AI - Cursorbench61.2Score63
FrontierMath - Tiers 1-3 (v2)65.61Accuracy (%, 285 private v2 problems)62.2
FrontierMath - Tier 4 (v2)29.27Accuracy (%, 41 private v2 problems)57.7
AutomationBench13.5Pass Rate (%)0

Interactive version: theaggregate.ai/model?slug=claude-sonnet-5-max · How the rankings work · Data refreshed daily, snapshot 2026-07-22.