Mistral Small 4 — benchmark results
March 2026 Mistral Small 4 model row. Provider: Mistral. Released 2026-03-16. Access: Open.
Unified ELO 1559 ± 22, rank #597 of 1776 rated models, from 46 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| SpeechMap Compliance | 90.9 | % Requests Completed | 91.4 |
| UGI - Writing | 40.33 | Writing Score | 74.5 |
| UGI - Natural Intelligence | 27.77 | NatInt Score | 69.4 |
| YapBench | 320.2 | YapIndex (lower is better) | 60.2 |
| Wolfram LLM Benchmarking Project | 43.1 | Correct Functionality (%) | 54.5 |
| AraGen | 33.85 | 3C3H Score (%) | 47.8 |
| ZeroEval GPQA Diamond | 71.2 | GPQA Diamond Score | 46.9 |
| UGI Leaderboard | 33.58 | UGI Score | 46.2 |
| AI for Education Pedagogy - Maths | 73.02 | Accuracy (%) | 43 |
| IDP Leaderboard | 69.6 | Avg Score (%) | 40 |
| RewardBench 2 Focus | 86.36 | Accuracy (%) | 39.2 |
| MATH-MC Level 4 | 96.25 | Accuracy (%) | 36 |
Interactive version: theaggregate.ai/model?slug=mistral-small-4 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.