Mistral Small 3.1 — benchmark results
Mistral Small 3.1 model row. Provider: Mistral. Released 2025-03-17. Access: Open.
Unified ELO 1473 ± 16, rank #895 of 1776 rated models, from 485 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval Italian Summarization - Ilpost SUM | 39.2 | Score (%) | 100 |
| MT-Bench PL - Coding | 8.3 | Judge Score (0-10) | 100 |
| EuroEval Catalan Summarization - Dacsa CA | 39.98 | Score (%) | 99.4 |
| EuroEval Dutch Summarization - Wiki Lingua NL | 36.35 | Score (%) | 99.3 |
| EuroEval Ukrainian Summarization - LR SUM UK | 32.26 | Score (%) | 99.3 |
| EuroEval Italian Common Sense Reasoning | 84.87 | Common Sense Reasoning Average Score (%) | 99.2 |
| EuroEval Albanian NLU - MMS SQ | 30.53 | Sentiment classification Score (%) | 99 |
| EuroEval German Common Sense Reasoning | 84.1 | Common Sense Reasoning Average Score (%) | 99 |
| EuroEval English Common Sense Reasoning | 90.82 | Common Sense Reasoning Average Score (%) | 98.7 |
| EuroEval Icelandic NLU - NQII | 63.53 | Reading comprehension Score (%) | 98.7 |
| EuroEval Icelandic Summarization - RRN | 39.67 | Score (%) | 98.7 |
| EuroEval Norwegian Summarization - NO Sammendrag | 31.99 | Score (%) | 98.7 |
Interactive version: theaggregate.ai/model?slug=mistral-small-3-1 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.