Mixtral 8x22B Instruct — benchmark results
Mistral's Apache-2.0 sparse-MoE instruct model (141B total/39B active, 64K context), strong at math and coding among April 2024 open models. Provider: Mistral. Released 2024-04-17. Access: Open.
Unified ELO 1559 ± 14, rank #596 of 1776 rated models, from 95 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AGC-Bench - grapheval_ai_researcher | 0.81 | Dataset z-score | 87.7 |
| AGC-Bench - tinystories | 0.94 | Dataset z-score | 85.4 |
| AGC-Bench - story_quality | 0.98 | Dataset z-score | 84.6 |
| AGC-Bench - future_ideas | 0.88 | Dataset z-score | 79.3 |
| AGC-Bench - schnovel | 0.9 | Dataset z-score | 78.7 |
| AGC-Bench - thenextchapter | 0.59 | Dataset z-score | 76.8 |
| AGC-Bench - grapheval_iclr | 0.65 | Dataset z-score | 75.6 |
| AGC-Bench - pollux_creativity | 0.75 | Dataset z-score | 75.6 |
| AGC-Bench - grapheval_review_advisor | 0.52 | Dataset z-score | 73.2 |
| AGC-Bench - conceptual_design | 0.61 | Dataset z-score | 70.7 |
| AGC-Bench - metaphoric_analogies | 0.32 | Dataset z-score | 70.7 |
| VNTL Leaderboard | 68.46 | Accuracy (%) | 67.4 |
Interactive version: theaggregate.ai/model?slug=mixtral-8x22b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.