Mistral Small: benchmark results
Mistral's small-tier workhorse under a rolling name, spanning the 22B (2409, research license) and 24B Apache-2.0 (2501) releases. Provider: Mistral. Released 2025-01-30. Access: Open.
Unified ELO 1493 ± 1, rank #723 of 1392 rated models, from 20 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| InfiBench | 55.62 | Score (%) | 87.6 |
| ProLLM - Entity Extraction | 81.8 | Score (%) | 58.5 |
| BenchmarkList ECI | 116.12 | Capability Index (ECI) | 56.3 |
| ProLLM - Function Calling | 87 | Score (%) | 52.8 |
| MixEval | 46.2 | Score | 52 |
| Wolfram LLM Benchmarking Project | 29.9 | Correct Functionality (%) | 33.7 |
| VAmoS Bench | 59.7 | Task Completion (%) | 33.3 |
| Klu LLM Leaderboard | 70 | Klu Index | 27.5 |
| ProLLM - SQL Disambiguation | 25.6 | Score (%) | 26.5 |
| GRIPS-hard | 14 | Accuracy (%) | 21.2 |
| Kagi LLM Benchmark | 37.8 | Accuracy (%) | 20.4 |
| ContextEcho | 27 | Delta (self-reported) | 18.2 |
Interactive version: theaggregate.ai/model?slug=mistral-small · How It Works · Data refreshed daily, snapshot 2026-09-05.