OLMo 3 7B Instruct — benchmark results

Allen AI OLMo 3 7B instruction-tuned checkpoint. Provider: Allen AI. Released 2025-11-20. Access: Open.

Unified ELO 1408 ± 10, rank #1182 of 1776 rated models, from 333 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EuroEval Portuguese NLU - ScaLA PT27.42Linguistic acceptability Score (%)82.5
EuroEval English NLU - ScaLA EN52.13Linguistic acceptability Score (%)78.4
EuroEval Italian NLU - ScaLA IT29.13Linguistic acceptability Score (%)76
EuroEval Polish Summarization - PSC22.96Score (%)70.7
EuroEval Dutch Summarization - Wiki Lingua NL33.57Score (%)70
EuroEval Catalan Summarization - Dacsa CA34.5Score (%)63.6
EuroEval French NLU - ScaLA FR37.65Linguistic acceptability Score (%)61.7
EuroEval Spanish NLU - ScaLA ES21.12Linguistic acceptability Score (%)59.9
Medmarks - MedHallu Hard45.94Score (%)58.6
EuroEval Catalan NLU - ScaLA CA12.47Linguistic acceptability Score (%)58.3
EuroEval Bosnian Summarization - LR SUM BS27.01Score (%)57.2
EuroEval Ukrainian Summarization - LR SUM UK26.96Score (%)56.6

Interactive version: theaggregate.ai/model?slug=olmo-3-7b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.