granite-3.0-8B-instruct — benchmark results
IBM's Apache-2.0 Granite 3.0 8B instruct model (October 2024), an enterprise workhorse tuned for RAG, summarization, extraction, and tool use. Provider: IBM. Released 2024-10-21. Access: Open.
Unified ELO 1459 ± 11, rank #960 of 1776 rated models, from 54 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - GPQA | 10.96 | Score | 81.4 |
| EuroEval Spanish NLU - Sentiment Headlines ES | 48.18 | Sentiment classification Score (%) | 81.3 |
| EuroEval Spanish NLU - MLQA ES | 64.53 | Reading comprehension Score (%) | 80.6 |
| EuroEval Swedish NLU - Swerec | 77.96 | Sentiment classification Score (%) | 80.6 |
| EuroEval Portuguese NLU - MultiWikiQA PT | 72.15 | Reading comprehension Score (%) | 68.4 |
| EuroEval Italian NLU - SQuAD IT | 69.69 | Reading comprehension Score (%) | 66.5 |
| EuroEval Danish NLU - Angry Tweets | 50.78 | Sentiment classification Score (%) | 63.9 |
| Open LLM Leaderboard - IFEval | 53.1 | Score | 63.1 |
| Open LLM Leaderboard - MATH Level 5 | 14.2 | Score | 60.6 |
| EuroEval Spanish NLU | 45.57 | NLU Average Score (%) | 59.8 |
| EuroEval Portuguese NLU - SST-2 PT | 78.64 | Sentiment classification Score (%) | 58.9 |
| Open LLM Leaderboard - BBH | 31.59 | Score | 57.6 |
Interactive version: theaggregate.ai/model?slug=granite-3-0-8b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.