salamandra-2B-instruct — benchmark results
Provider: BSC. Released 2024-09-30. Access: Open.
Unified ELO 1271 ± 45, rank #1587 of 1776 rated models, from 19 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LA Leaderboard - XNLI Galician | 46.84 | Accuracy (%) | 75 |
| LA Leaderboard - PIQA Catalan | 64.31 | Accuracy (%) | 69.1 |
| EuroEval Portuguese NLU - SST-2 PT | 73.31 | Sentiment classification Score (%) | 43 |
| EuroEval Finnish NLU - Scandisent FI | 85.18 | Sentiment classification Score (%) | 33.2 |
| LA Leaderboard - AQuAS | 62.14 | Accuracy (%) | 30.9 |
| EuroEval Icelandic Knowledge | 2.48 | Knowledge Average Score (%) | 28.2 |
| EuroEval Dutch NLU - DBRD | 64.59 | Sentiment classification Score (%) | 24.5 |
| EuroEval Norwegian Knowledge | 2.94 | Knowledge Average Score (%) | 21.2 |
| EuroEval Portuguese NLU - ScaLA PT | 1.14 | Linguistic acceptability Score (%) | 20.7 |
| EuroEval Italian NLU - Sentipolc16 | 29.8 | Sentiment classification Score (%) | 20.1 |
| LA Leaderboard | 48.81 | Average Score (%) | 19.1 |
| LA Leaderboard - COPA Spanish | 69.4 | Accuracy (%) | 19.1 |
Interactive version: theaggregate.ai/model?slug=salamandra-2b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.