Mistral-7B-OpenOrca — benchmark results
Provider: Mistral. Released 2023-09-29. Access: Open.
Unified ELO 1261 ± 27, rank #1596 of 1776 rated models, from 16 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| LLM Trustworthy - Out-of-Distribution | 73.4 | Trust Score (%) | 76 |
| LLM Trustworthy - Adversarial | 47.23 | Trust Score (%) | 60 |
| Open LLM Leaderboard - IFEval | 49.78 | Score | 58.5 |
| Open Korean LLM Leaderboard | 193.84 | Average Score (%) | 58.4 |
| LLM Trustworthy - Adversarial Demo | 62.15 | Trust Score (%) | 52 |
| Open LLM Leaderboard - BBH | 25.84 | Score | 39.9 |
| LLM Trustworthy - Privacy | 77.36 | Trust Score (%) | 36 |
| Open LLM Leaderboard - MuSR | 5.89 | Score | 28.3 |
| Open LLM Leaderboard - MMLU-Pro | 18.37 | Score | 28 |
| Open LLM Leaderboard - GPQA | 2.91 | Score | 26.7 |
| LLM Trustworthy - Stereotype | 79.33 | Trust Score (%) | 24 |
| LLM Trustworthy - Toxicity | 30.12 | Trust Score (%) | 24 |
Interactive version: theaggregate.ai/model?slug=mistral-7b-openorca · How the rankings work · Data refreshed daily, snapshot 2026-07-22.