Codestral-22B-v0.1 — benchmark results
Mistral's 22B code-generation model (May 2024) covering 80+ programming languages with fill-in-the-middle, under the non-production MNPL license. Provider: Mistral. Released 2024-05-29. Access: Open.
Unified ELO 1416 ± 26, rank #1152 of 1776 rated models, from 14 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EvalPlus (HumanEval+ & MBPP+) | 67.8 | Pass@1 avg (%) | 75.8 |
| Open LLM Leaderboard - IFEval | 57.72 | Score | 69.7 |
| BigCodeBench | 41.8 | Pass@1 (%) | 64.3 |
| DuckDB-NSQL | 54.7 | Execution Accuracy (%) | 57.3 |
| Open LLM Leaderboard - MuSR | 10.74 | Score | 54.6 |
| Open LLM Leaderboard - GPQA | 6.49 | Score | 54.5 |
| Open LLM Leaderboard - BBH | 30.74 | Score | 54.4 |
| CodeElo | 385 | Elo Rating | 52.9 |
| Open LLM Leaderboard - MATH Level 5 | 10.05 | Score | 47.6 |
| Open LLM Leaderboard - MMLU-Pro | 23.95 | Score | 42.3 |
| LiveOIBench | 15.82 | Avg Human Percentile | 22.8 |
| IFEval Leaderboard | 60.74 | Final Score | 20 |
Interactive version: theaggregate.ai/model?slug=codestral-22b-v0-1 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.