Yi-1.5-6B-Chat: benchmark results
Provider: 01.AI. Released 2024-05-11. Access: Open.
Unified ELO 1395 ± 1, rank #1157 of 1392 rated models, from 61 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 14.03 | Score | 76.9 |
| LatamBoard - Spanish PAWS | 61.5 | Score (%) | 72.7 |
| Open LLM Leaderboard - MATH Level 5 | 16.24 | Score | 64.2 |
| LatamBoard - Spanish MGSM | 9.6 | Score (%) | 63.6 |
| Open LLM Leaderboard - IFEval | 51.45 | Score | 61.3 |
| Open LLM Leaderboard - GPQA | 6.94 | Score | 58.3 |
| Mobile-MMLU | 60.5 | Overall Accuracy (%) | 57.1 |
| MAGI-Hard | 46.18 | Accuracy (%, 2024-05 snapshot) | 53.4 |
| MERA - BPS | 95.7 | Accuracy (%) | 51.9 |
| LatamBoard - Spanish Escola | 66.1 | Score (%) | 50 |
| Open LLM Leaderboard - MMLU-Pro | 24.37 | Score | 43.6 |
| MERA - RWSD | 54.62 | Accuracy (%) | 40.9 |
Interactive version: theaggregate.ai/model?slug=yi-1-5-6b-chat · How It Works · Data refreshed daily, snapshot 2026-09-05.