gpt-sw3-126m-instruct — benchmark results

Provider: AI Sweden. Released 2023-04-28. Access: Open.

Unified ELO 1019 ± 43, rank #1772 of 1776 rated models, from 94 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
EuroEval Icelandic NLU - ScaLA IS2.31Linguistic acceptability Score (%)45.8
EuroEval Icelandic Common Sense Reasoning0.59Common Sense Reasoning Average Score (%)19.8
EuroEval Icelandic NLU - MIM-GOLD NER17.85Named entity recognition Score (%)18.7
Icelandic LLM - WinoGrande-IS53.86Score (%)18.7
Icelandic LLM - GED51Score (%)17
EuroEval Danish NLU - Angry Tweets17.37Sentiment classification Score (%)16.3
EuroEval Faroese NLU - ScaLA FO0Linguistic acceptability Score (%)15.9
EuroEval Norwegian8.13Average Score (%)15.8
EuroEval Dutch Common Sense Reasoning0.69Common Sense Reasoning Average Score (%)14.8
EuroEval Faroese NLU - FONE29.27Named entity recognition Score (%)14.7
EuroEval Norwegian NLU - Norquad9.58Reading comprehension Score (%)14.4
EuroEval Icelandic NLU - Hotter and Colder Sentiment3.39Sentiment classification Score (%)14

Interactive version: theaggregate.ai/model?slug=gpt-sw3-126m-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.