Magistral Medium 1.2: benchmark results
Provider: Mistral. Released 2025-09-17. Access: API.
Unified ELO 1684 ± 19, rank #358 of 2656 rated models, from 107 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA LiveCodeBench | 75.03 | Pass@1 (%) | 86.5 |
| CritPt | 30 | Accuracy (self-reported) | 80.8 |
| SnorkelSpatial | 44.24 | Accuracy@1 (%, 330 spatial reasoning questions) | 78.1 |
| AA AIME 2025 | 82 | Accuracy (%) | 77.6 |
| AA-Omniscience Hallucination Rate | 59.98 | Hallucination Rate (%) | 75.9 |
| AA MMLU-Pro | 81.51 | Accuracy (%) | 75.3 |
| AIME 2024 | 91.8 | Score (percentage) | 75 |
| Wolfram LLM Benchmarking Project | 51.7 | Correct Functionality (%) | 72.1 |
| AA Omniscience - Software Engineering (SWE) - Rust | 54 | Accuracy (%) | 68.9 |
| AA Omniscience - Software Engineering (SWE) - Julia | 16 | Accuracy (%) | 68.8 |
| AA Omniscience - Software Engineering (SWE) - Kotlin | 22 | Accuracy (%) | 68.6 |
| AA Omniscience - Software Engineering (SWE) - Go | 22 | Accuracy (%) | 68.4 |
Interactive version: theaggregate.ai/model?slug=magistral-medium-1-2 · How It Works · Data refreshed daily, snapshot 2026-09-19.