Mistral Large — benchmark results
February 2024 Mistral Large snapshot row. Provider: Mistral. Released 2024-02-26. Access: API.
Unified ELO 1497 ± 14, rank #809 of 1776 rated models, from 95 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AgentLeak | 99.3 | Total Leak (self-reported) | 100 |
| BlueBench - Safety | 88.04 | Score (%) | 94.1 |
| InfiBench | 58.22 | Score (%) | 92.4 |
| TextClass Benchmark | 1720.36 | Meta-Elo (self-reported) | 88.8 |
| BiGGen-Bench | 3.93 | Average Score (1-5) | 88.2 |
| SEAL - Adversarial Robustness | 37 | Score | 85.7 |
| LiveBench Zebra Puzzle | 40 | Score | 85.4 |
| LiveBench Summarize | 69.23 | Score | 83.3 |
| LiveBench Table Reformat | 64 | Score | 83.3 |
| BlueBench - Bias | 96.97 | Score (%) | 82.4 |
| BlueBench - Summarization | 18.24 | Score (%) | 82.4 |
| CZ-EVAL - Analytical | 38.52 | Accuracy (%) | 81.8 |
Interactive version: theaggregate.ai/model?slug=mistral-large · How the rankings work · Data refreshed daily, snapshot 2026-07-22.