Yi-1.5-6B — benchmark results
01.AI's Apache-2.0 6B base model (May 2024): Yi-6B continued-pretrained on 500B additional tokens for stronger coding, math, and reasoning. Provider: 01.AI. Released 2024-05-11. Access: Open.
Unified ELO 1414 ± 16, rank #1160 of 1776 rated models, from 34 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 13.31 | Score | 73.4 |
| EuroEval Portuguese NLU - MultiWikiQA PT | 72.84 | Reading comprehension Score (%) | 72.6 |
| Open Chinese LLM - WinoGrande | 65.04 | Accuracy (%) | 70.8 |
| Open LLM Leaderboard - GPQA | 8.5 | Score | 70.7 |
| Open Chinese LLM - CMMLU | 57.54 | Accuracy (%) | 64.7 |
| Open Chinese LLM - C-Eval Semantic | 72.2 | Accuracy (%) | 58.8 |
| EuroEval Dutch NLU - DBRD | 87.76 | Sentiment classification Score (%) | 54.3 |
| EuroEval Portuguese NLU - SST-2 PT | 77.09 | Sentiment classification Score (%) | 51.1 |
| Open Chinese LLM - HellaSwag | 60.24 | Accuracy (%) | 50.4 |
| EuroEval Portuguese NLU | 47.47 | NLU Average Score (%) | 44.6 |
| Open Chinese LLM Leaderboard | 56.04 | Average Score (%) | 43.9 |
| EuroEval English Knowledge | 69.09 | Knowledge Average Score (%) | 43.8 |
Interactive version: theaggregate.ai/model?slug=yi-1-5-6b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.