salamandra-2B — benchmark results

Provider: BSC. Released 2024-09-30. Access: Open.

Unified ELO 1258 ± 52, rank #1604 of 1776 rated models, from 19 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LA Leaderboard - XNLI Galician45.9Accuracy (%)58.8
LA Leaderboard - GalCoLA51.61Accuracy (%)56.6
LA Leaderboard - PIQA Catalan62.02Accuracy (%)53.7
EuroEval Finnish NLU - Scandisent FI88.87Sentiment classification Score (%)48.3
EuroEval Portuguese NLU - SST-2 PT71.59Sentiment classification Score (%)37.7
LA Leaderboard - COPA Spanish71.2Accuracy (%)27.2
EuroEval Dutch NLU - DBRD69.02Sentiment classification Score (%)26.7
EuroEval Italian NLU - ScaLA IT2.88Linguistic acceptability Score (%)24.2
EuroEval Italian NLU - Sentipolc1629.64Sentiment classification Score (%)19.6
LA Leaderboard - Spanish Law Exams25.21Accuracy (%)16.2
LA Leaderboard48.27Average Score (%)14.7
EuroEval Danish NLU - Angry Tweets15.68Sentiment classification Score (%)13.8

Interactive version: theaggregate.ai/model?slug=salamandra-2b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.