Claude 3.5 Sonnet (20240620) — benchmark results
June 2024 Claude 3.5 Sonnet snapshot, kept separate from later Claude 3.5 Sonnet rows. Provider: Anthropic. Released 2024-06-20. Access: API.
Unified ELO 1670 ± 9, rank #320 of 1776 rated models, from 687 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HELM EWoK - Material Dynamics | 90.9 | EM | 100 |
| HELM EWoK - Material Properties | 93.95 | EM | 100 |
| HELM EWoK - Physical Relations | 83.61 | EM | 100 |
| HELM NaturalQuestions (Closed) | 50.16 | F1 (%) | 100 |
| HELM ThaiExam - IC | 87.37 | EM | 100 |
| HELM ThaiExam - ONET | 73.46 | EM | 100 |
| HELM ThaiExam - TGAT | 83.08 | EM | 100 |
| HELM ThaiExam - ThaiExam | 75.1 | EM | 100 |
| LLM Game Benchmark | 42.2 | Source Win Rate (self-reported) | 100 |
| LiveBench AMPS Hard | 49 | Score | 100 |
| LiveBench Coding Completion | 68.42 | Score | 100 |
| LiveBench LCB Generation | 58 | Score | 100 |
Interactive version: theaggregate.ai/model?slug=claude-3-5-sonnet-20240620 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.