gpt-sw3-20B-instruct — benchmark results
Provider: AI Sweden. Released 2023-04-28. Access: Open.
Unified ELO 1313 ± 35, rank #1505 of 1776 rated models, from 54 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Icelandic NLU - NQII | 53.88 | Reading comprehension Score (%) | 81.5 |
| EuroEval English NLU - SQuAD | 83.08 | Reading comprehension Score (%) | 74.2 |
| EuroEval Norwegian NLU - Norquad | 67.93 | Reading comprehension Score (%) | 64.4 |
| EuroEval English NLU - SST-5 | 65.58 | Sentiment classification Score (%) | 56.9 |
| EuroEval Icelandic | 25.68 | Average Score (%) | 45.7 |
| EuroEval Icelandic NLU - ScaLA IS | 1.95 | Linguistic acceptability Score (%) | 44.1 |
| EuroEval Norwegian NLU - Norec | 40.89 | Sentiment classification Score (%) | 43.1 |
| EuroEval Norwegian NLU - ScaLA NN | 8.32 | Linguistic acceptability Score (%) | 37.6 |
| EuroEval Icelandic NLU | 25.68 | NLU Average Score (%) | 35.8 |
| EuroEval Swedish NLU - Swerec | 68.23 | Sentiment classification Score (%) | 33.5 |
| EuroEval Swedish NLU - ScaLA SV | 12.39 | Linguistic acceptability Score (%) | 33 |
| EuroEval Norwegian NLU - ScaLA NB | 9.45 | Linguistic acceptability Score (%) | 32.1 |
Interactive version: theaggregate.ai/model?slug=gpt-sw3-20b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.