Mistral Large 3 — benchmark results
Mistral Large 3 model row. Provider: Mistral. Released 2025-12-02. Access: API.
Unified ELO 1606 ± 12, rank #466 of 1776 rated models, from 274 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LLM Stats (MM-MT-Bench) | 84.9 | Score (%) | 100 |
| ROK-FORTRESS | 28.4 | Î_ling (pp) (self-reported) | 100 |
| XL-SafetyBench | 98.8 | Overall ASR (self-reported) | 100 |
| SpeechMap Compliance | 98.2 | % Requests Completed | 98.9 |
| UGI Leaderboard | 56.86 | UGI Score | 97.7 |
| AGC-Bench - ocw | 1.63 | Dataset z-score | 97.6 |
| AGC-Bench - cpers | 1.14 | Dataset z-score | 93.9 |
| AGC-Bench - mops | 1.46 | Dataset z-score | 93.9 |
| LLM Stats (Wild Bench) | 68.5 | Score (%) | 92.9 |
| AGC-Bench - thenextchapter | 1.13 | Dataset z-score | 90.2 |
| UGI - Natural Intelligence | 39.98 | NatInt Score | 87.2 |
| HumainE Leaderboard | 32.68 | HumainE Score | 86.8 |
Interactive version: theaggregate.ai/model?slug=mistral-large-3 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.