Magistral Medium: benchmark results
Mistral's first frontier-class reasoning model, the larger API-only counterpart to Magistral Small (June 2025). Provider: Mistral. Released 2025-06-10. Access: API.
Unified ELO 1527 ± 1, rank #556 of 1392 rated models, from 75 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| SpeechMap Compliance | 92.6 | % Requests Completed | 94.3 |
| CritPt | 30 | Accuracy (self-reported) | 80.8 |
| HumainE Leaderboard | 30.88 | HumainE Score | 73.6 |
| SnakeBench | 25.8 | TrueSkill Rating | 70.2 |
| BenchTable | 55.9 | Total Score (%) | 63.5 |
| AA Omniscience - Science, Engineering & Mathematics | 31 | Accuracy (%) | 60.7 |
| AI for Education Pedagogy - Social studies | 80.91 | Accuracy (%) | 59.5 |
| AA Omniscience | -26.93 | Score | 58.9 |
| AA Omniscience - Health | 22.2 | Accuracy (%) | 56.9 |
| AA Omniscience - Business | 18.1 | Accuracy (%) | 56.4 |
| AA Humanity's Last Exam | 9.82 | Accuracy (%) | 53.4 |
| AA CritPt | 0.29 | Accuracy (%) | 53.2 |
Interactive version: theaggregate.ai/model?slug=magistral-medium · How It Works · Data refreshed daily, snapshot 2026-09-05.