Mistral 7B Instruct (v0.1) — benchmark results
Provider: Mistral. Released 2023-09-27. Access: Open.
Unified ELO 1303 ± 23, rank #1526 of 1776 rated models, from 100 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| NPHardEval - GCP | 20.91 | Accuracy (%) | 100 |
| NPHardEval - GCP D | 58 | Accuracy (%) | 90.9 |
| NPHardEval - TSP D | 62.73 | Accuracy (%) | 81.8 |
| Open Medical LLM - PubMedQA | 75.8 | Accuracy (%) | 78.4 |
| FastEval | 42.66 | Total Score | 62.5 |
| SpeechMap Compliance | 70.4 | % Requests Completed | 62.1 |
| PIQA | 82.2 | Accuracy (%) | 60.8 |
| EuroEval Portuguese NLU - MultiWikiQA PT | 70.47 | Reading comprehension Score (%) | 60.1 |
| EuroEval Spanish NLU - Sentiment Headlines ES | 41.35 | Sentiment classification Score (%) | 56.6 |
| EuroEval Italian NLU - Sentipolc16 | 50.71 | Sentiment classification Score (%) | 52.9 |
| EuroEval Spanish NLU - MLQA ES | 59.98 | Reading comprehension Score (%) | 51.8 |
| CyberBench (NLP) | 58.38 | Avg Score (%) | 50 |
Interactive version: theaggregate.ai/model?slug=mistral-7b-instruct-v0-1 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.