mamba-2.8B-hf — benchmark results
Provider: Other. Released 2024-03-05. Access: Open.
Unified ELO 1249 ± 20, rank #1620 of 1776 rated models, from 11 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Dutch NLU - DBRD | 69.95 | Sentiment classification Score (%) | 27.4 |
| EuroEval Portuguese NLU - SST-2 PT | 58.23 | Sentiment classification Score (%) | 22.7 |
| EuroEval Portuguese NLU - ScaLA PT | 1.62 | Linguistic acceptability Score (%) | 22.7 |
| EuroEval Dutch NLU - SQuAD NL | 34.26 | Reading comprehension Score (%) | 21 |
| EuroEval Portuguese NLU - MultiWikiQA PT | 26.85 | Reading comprehension Score (%) | 19.6 |
| EuroEval Portuguese Common Sense Reasoning | 0.87 | Common Sense Reasoning Average Score (%) | 15.1 |
| EuroEval Portuguese NLU | 25.84 | NLU Average Score (%) | 15 |
| EuroEval Portuguese | 17.28 | Average Score (%) | 12.2 |
| BABILong (NIAH) | 33.2 | Avg Accuracy (%) | 10.3 |
| EuroEval Portuguese NLU - HAREM | 16.66 | Named entity recognition Score (%) | 9.2 |
| EuroEval Portuguese Knowledge | -0.55 | Knowledge Average Score (%) | 4.6 |
Interactive version: theaggregate.ai/model?slug=mamba-2-8b-hf · How the rankings work · Data refreshed daily, snapshot 2026-07-22.