EuroLLM-22B-Instruct-2512 — benchmark results

The EU-funded EuroLLM project's 22B instruct model covering the 24 official EU languages plus 11 others, trained on EuroHPC compute (December 2025). Provider: EuroLLM. Released 2025-12-01. Access: Open.

Unified ELO 1498 ± 15, rank #803 of 1776 rated models, from 276 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EuroEval Croatian NLU - MMS HR46.07Sentiment classification Score (%)96.7
EuroEval German NLU - ScaLA DE65.65Linguistic acceptability Score (%)96.6
EuroEval Greek NLU - ScaLA EL57.12Linguistic acceptability Score (%)96.1
EuroEval Italian NLU - ScaLA IT59.27Linguistic acceptability Score (%)96.1
EuroEval Estonian NLU - Grammar ET51.64Linguistic acceptability Score (%)96
EuroEval Portuguese NLU - ScaLA PT55.09Linguistic acceptability Score (%)95
EuroEval Spanish NLU - Sentiment Headlines ES50.66Sentiment classification Score (%)92.7
EuroEval Danish Summarization - Nordjylland News37.44Score (%)92.4
EuroEval Estonian NLU - Estonian Valence62.13Sentiment classification Score (%)92
EuroEval Polish NLU - ScaLA PL55.16Linguistic acceptability Score (%)91.4
EuroEval German Summarization - Mlsum DE37.54Score (%)91.2
EuroEval Norwegian Summarization - NO Sammendrag30.98Score (%)90.9

Interactive version: theaggregate.ai/model?slug=eurollm-22b-instruct-2512 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.