Mistral Large 2512: benchmark results
Provider: Mistral. Released 2025-12-02. Access: Open.
Unified ELO 1719 ± 29, rank #291 of 2656 rated models, from 14 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| TaxEval v2 | 73.06 | Score (%) | 64.9 |
| Vals AI CaseLaw v2 | 61.41 | Accuracy (%) | 61.9 |
| CorpFin v2 | 61.03 | Score (%) | 56.8 |
| LegalBench | 79.14 | Score (%) | 37.5 |
| MMLU Pro | 79.82 | Score (%) | 35.5 |
| AIME | 42.92 | Score (%) | 31.6 |
| GPQA Diamond | 68.43 | Score (%) | 30.4 |
| MedQA | 82.23 | Score (%) | 29.8 |
| MMMU Pro | 66.19 | Score (%) | 22.4 |
| IOI | 4 | Score (self-reported) | 20 |
| MortgageTax | 52.11 | Score (%) | 16.9 |
| LLM-SoccerArena | 144 | Prediction Points | 16.7 |
Interactive version: theaggregate.ai/model?slug=mistral-large-2512 · How It Works · Data refreshed daily, snapshot 2026-09-19.