mistral-small-latest — benchmark results
Rolling alias for Mistral's current Small - as of mid-2026 Mistral Small 4, a 119B MoE (6.5B active) hybrid unifying chat, reasoning and coding (256K context). Provider: Mistral. Released 2026-03-16. Access: Open.
Unified ELO 1519 ± 28, rank #732 of 1776 rated models, from 9 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Arcadia MMMU Open Ended | 35.85 | Accuracy (%) | 90 |
| Turkish MMLU | 67 | Accuracy (%) | 69.2 |
| CyberSecEval2 Prompt Injection | 28.59 | Accuracy (%) | 44.4 |
| Arcadia MMMU Multiple Choice | 54.55 | Accuracy (%) | 36.4 |
| OpenAI HumanEval | 82.32 | Accuracy (%) | 31.8 |
| CyberSecEval2 Interpreter Abuse | 18.55 | Accuracy (%) | 22.2 |
| Arcadia CommonsenseQA | 82.23 | Accuracy (%) | 10 |
| CyberSecEval2 Vulnerability Exploit | 35.91 | Accuracy (%) | 10 |
| GDM InterCode CTF | 35.44 | Accuracy (%) | 0 |
Interactive version: theaggregate.ai/model?slug=mistral-small-latest · How the rankings work · Data refreshed daily, snapshot 2026-07-22.