Claude Sonnet 5.5 (Max): benchmark results

Provider: Anthropic. Access: API.

Unified ELO 1954 ± 24, rank #18 of 2055 rated models, from 33 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Bug Hunt Bench57Planted Bugs Fixed (out of 105)100
Bug Hunt Bench - LMS33Planted Bugs Fixed (out of 60)100
Bug Hunt Bench - VS Code Extension24Planted Bugs Fixed (out of 45)100
LiveBench Code Generation95.78Score100
CursorBench 4.055.5Score (%)98.3
Vals AI MedScribe91.1Accuracy (%)98.1
LiveBench Code Completion86.96Score96.8
LiveBench Theory of Mind86.54Score94.4
LiveBench Olympiad92.47Score93.7
LiveBench JavaScript77.27Score92.9
LiveBench TypeScript56.67Score91.3
Vals AI MedCode52.92Accuracy (%)90.1

Interactive version: theaggregate.ai/model?slug=claude-sonnet-5-5-max · How It Works · Data refreshed daily, snapshot 2026-09-29.