gpt-sw3-20B: benchmark results
Provider: AI Sweden. Released 2022-12-14. Access: Open.
Unified ELO 1250 ± 69, rank #2575 of 2656 rated models, from 43 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Danish NLU - Multi Wiki QA DA | 72.47 | Reading comprehension Score (%) | 57.6 |
| EuroEval Swedish NLU - Multi Wiki QA SV | 67.19 | Reading comprehension Score (%) | 48.2 |
| EuroEval Norwegian Knowledge - Idioms NO | 2.73 | MCC (x100) | 45.6 |
| EuroEval Swedish NLU - ScaLA SV | 12.95 | Linguistic acceptability Score (%) | 33.7 |
| EuroEval Danish NLU - ScaLA DA | 8.38 | Linguistic acceptability Score (%) | 31.8 |
| EuroEval Faroese NLU - FoSent | 8.4 | Sentiment classification Score (%) | 29.8 |
| EuroEval Norwegian Knowledge | 4.61 | Knowledge Average Score (%) | 27.5 |
| EuroEval Danish NLU | 26.95 | NLU Average Score (%) | 27 |
| EuroEval Danish Knowledge - Danish Citizen Tests | 19.87 | MCC (x100) | 25.8 |
| EuroEval Faroese NLU - ScaLA FO | 0.44 | Linguistic acceptability Score (%) | 24.4 |
| EuroEval Icelandic Knowledge | 1.67 | Knowledge Average Score (%) | 24.2 |
| Open Portuguese LLM - OAB Exams | 27.79 | Accuracy (%) | 22.8 |
Interactive version: theaggregate.ai/model?slug=gpt-sw3-20b · How It Works · Data refreshed daily, snapshot 2026-09-19.