Claude 3.5 Haiku (20241022) — benchmark results

October 22, 2024 Claude 3.5 Haiku snapshot, tracked when sources report the dated API model. Provider: Anthropic. Released 2024-10-22. Access: API.

Unified ELO 1450 ± 19, rank #997 of 1776 rated models, from 196 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EuroEval Lithuanian NLU - ScaLA LT50.62Linguistic acceptability Score (%)95.6
EuroEval Faroese NLU - FoSent65.97Sentiment classification Score (%)92.2
EuroEval Finnish NLU - Scandisent FI92.47Sentiment classification Score (%)92.2
AILuminate Safety4Safety Grade (1-5)91.9
EuroEval Polish NLU - ScaLA PL53.53Linguistic acceptability Score (%)89.3
UGI - Natural Intelligence37.98NatInt Score86.5
EuroEval Latvian NLU - Latvian Twitter Sentiment49.13Sentiment classification Score (%)84.9
BigCodeBench46.1Pass@1 (%)84.1
Galileo Tool Tasks - xLAM Tool Missing76Accuracy (%)83.3
EuroEval Faroese NLU - ScaLA FO19.17Linguistic acceptability Score (%)83.2
EuroEval Norwegian NLU - ScaLA NB60.8Linguistic acceptability Score (%)83.1
EuroEval Latvian NLU - ScaLA LV33.1Linguistic acceptability Score (%)81.9

Interactive version: theaggregate.ai/model?slug=claude-3-5-haiku-20241022 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.