gpt-sw3-1.3B-instruct — benchmark results
Provider: AI Sweden. Released 2023-04-28. Access: Open.
Unified ELO 1241 ± 89, rank #1639 of 1776 rated models, from 18 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Swedish NLU - Swerec | 73.34 | Sentiment classification Score (%) | 42.5 |
| EuroEval Norwegian NLU - Norec | 35.58 | Sentiment classification Score (%) | 36.8 |
| EuroEval Danish NLU - Angry Tweets | 27.14 | Sentiment classification Score (%) | 24.6 |
| EuroEval Danish | 15.44 | Average Score (%) | 16.8 |
| EuroEval Norwegian | 8.88 | Average Score (%) | 16.6 |
| EuroEval Danish NLU - Multi Wiki QA DA | 17.3 | Reading comprehension Score (%) | 15.2 |
| EuroEval Swedish NLU - Multi Wiki QA SV | 15.01 | Reading comprehension Score (%) | 14.2 |
| Icelandic LLM - WinoGrande-IS | 52.11 | Score (%) | 13.7 |
| EuroEval Norwegian Knowledge | 1.29 | Knowledge Average Score (%) | 13.3 |
| Icelandic LLM - GED | 50.5 | Score (%) | 13.2 |
| Icelandic LLM - WikiQA-IS | 2.84 | Score (%) | 11 |
| Icelandic LLM - Belebele-IS | 22.89 | Score (%) | 8.8 |
Interactive version: theaggregate.ai/model?slug=gpt-sw3-1-3b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.