Claude Haiku 5.5: benchmark results

Provider: Anthropic. Access: API.

Unified ELO 1845 ± 18, rank #30 of 1632 rated models, from 133 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LiveBench Math Comp98.04Score98.5
AA-Omniscience Index - Software Engineering (SWE) - Rust84Omniscience Index97.9
LLM Stats Score50.31LLM Stats Score (conservative rating)95.5
AA-Briefcase - Analytical Quality Elo1907.63Elo94.6
AA GDP.pdf - HR - Criterion Pass Rate84.12Criterion Pass Rate (%)94.3
AA Long Context Reasoning82.67Accuracy (%)94.1
Artificial Analysis Intelligence Index43.4Intelligence Index93.9
AA Humanity's Last Exam44.39Accuracy (%)92.9
LiveBench Theory of Mind86.54Score92.6
AutomationBench-AA - Operations - Raw Objective Completion95.08Raw Objective Completion (%)92.4
AA-Omniscience Index - Software Engineering (SWE) - JavaScript70Omniscience Index92.3
AA-Omniscience Index - Software Engineering (SWE) - Swift64Omniscience Index92.3

Interactive version: theaggregate.ai/model?slug=claude-haiku-5-5 · How It Works · Data refreshed daily, snapshot 2026-10-08.