granite-3.2-8B-instruct — benchmark results
IBM's Granite 3.2 8B instruct model with an experimental toggleable thinking mode for step-by-step reasoning. Provider: IBM. Released 2025-02-26. Access: Open.
Unified ELO 1514 ± 25, rank #754 of 1776 rated models, from 8 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 16.79 | Score | 86.8 |
| Open LLM Leaderboard - IFEval | 72.75 | Score | 86.7 |
| Open LLM Leaderboard - MATH Level 5 | 23.79 | Score | 77.2 |
| Open LLM Leaderboard - GPQA | 8.72 | Score | 71.9 |
| Open LLM Leaderboard - BBH | 34.66 | Score | 69.1 |
| Open LLM Leaderboard - MMLU-Pro | 27.91 | Score | 51.8 |
| Gorilla API Bench (BFCL) | 26.87 | Overall Accuracy (%) | 25.6 |
| Wolfram LLM Benchmarking Project | 23 | Correct Functionality (%) | 21.7 |
Interactive version: theaggregate.ai/model?slug=granite-3-2-8b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.