Magistral Medium — benchmark results
Mistral's first frontier-class reasoning model, the larger API-only counterpart to Magistral Small (June 2025). Provider: Mistral. Released 2025-06-10. Access: API.
Unified ELO 1553 ± 19, rank #613 of 1776 rated models, from 85 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| SpeechMap Compliance | 92.6 | % Requests Completed | 93.8 |
| HumainE Leaderboard | 30.88 | HumainE Score | 73.6 |
| AA MATH-500 | 91.73 | Accuracy (%) | 68.8 |
| AA Omniscience - Software Engineering (SWE) - Rust | 58 | Accuracy (%) | 68.8 |
| AA Omniscience - Science, Engineering & Mathematics | 30.9 | Accuracy (%) | 68.4 |
| AA Omniscience | -26.42 | Score | 66.5 |
| AA Omniscience - Software Engineering (SWE) - Swift | 44 | Accuracy (%) | 63.3 |
| AI for Education Pedagogy - Social studies | 80.91 | Accuracy (%) | 63.2 |
| AA Omniscience - Health | 22.1 | Accuracy (%) | 62.5 |
| AA Omniscience - Business | 19.2 | Accuracy (%) | 61.5 |
| AA Humanity's Last Exam | 9.55 | Accuracy (%) | 59.9 |
| AA LiveCodeBench | 52.7 | Pass@1 (%) | 59.4 |
Interactive version: theaggregate.ai/model?slug=magistral-medium · How the rankings work · Data refreshed daily, snapshot 2026-07-22.