Claude Haiku 4.5 (20251001) (Thinking): benchmark results
Claude Haiku 4.5 (20251001) evaluated with thinking enabled. Provider: Anthropic. Released 2025-10-01. Access: API.
Unified ELO 1542 ± 1, rank #663 of 1761 rated models, from 38 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| UGI - Writing | 49.08 | Writing Score | 87.9 |
| UGI - Natural Intelligence | 36.76 | NatInt Score | 85.1 |
| Vals AI MedScribe | 85.23 | Accuracy (%) | 76.8 |
| Vals AI MGSM | 92.15 | Accuracy (%) | 72.7 |
| Vals AI AIME | 82.71 | Accuracy (%) | 54.7 |
| Vals AI CorpFin v2 | 60.61 | Accuracy (%) | 49.6 |
| Vals AI LegalBench | 81.24 | Accuracy (%) | 46.8 |
| Vals AI Terminal-Bench 2.0 | 38.2 | Accuracy (%) | 46.2 |
| Vals AI CaseLaw v2 | 56.48 | Accuracy (%) | 44.1 |
| Vals AI MortgageTax | 62.16 | Accuracy (%) | 43.3 |
| Vals AI Finance Agent | 46.93 | Accuracy (%) | 40 |
| Vals AI GPQA | 72.22 | Accuracy (%) | 34.7 |
Interactive version: theaggregate.ai/model?slug=claude-haiku-4-5-20251001-thinking · How It Works · Data refreshed daily, snapshot 2026-09-05.