granite-3.0-3B-a800m-instruct — benchmark results
Provider: IBM. Released 2024-10-03. Access: Open.
Unified ELO 1340 ± 24, rank #1437 of 1776 rated models, from 37 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Spanish NLU - MLQA ES | 62.81 | Reading comprehension Score (%) | 68.1 |
| EuroEval Portuguese NLU - SST-2 PT | 79.46 | Sentiment classification Score (%) | 62.7 |
| EuroEval Dutch NLU - DBRD | 88.43 | Sentiment classification Score (%) | 58.7 |
| EuroEval Italian NLU - SQuAD IT | 65.67 | Reading comprehension Score (%) | 49.4 |
| Open LLM Leaderboard - IFEval | 42.98 | Score | 45 |
| EuroEval Spanish NLU - ScaLA ES | 4.87 | Linguistic acceptability Score (%) | 39.3 |
| EuroEval Italian NLU - Sentipolc16 | 44.07 | Sentiment classification Score (%) | 37.6 |
| Open LLM Leaderboard - MATH Level 5 | 7.02 | Score | 37.5 |
| Open LLM Leaderboard - GPQA | 4.14 | Score | 36 |
| EuroEval Spanish NLU | 36.18 | NLU Average Score (%) | 34.4 |
| EuroEval Portuguese NLU - ScaLA PT | 4.32 | Linguistic acceptability Score (%) | 33.4 |
| EuroEval Spanish Common Sense Reasoning | 19.35 | Common Sense Reasoning Average Score (%) | 32.9 |
Interactive version: theaggregate.ai/model?slug=granite-3-0-3b-a800m-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.