Yi Large — benchmark results
01.AI's closed API flagship chat model with strong multilingual performance, the Chinese startup's top proprietary offering of 2024. Provider: 01.AI. Released 2024-06-25. Access: API.
Unified ELO 1519 ± 15, rank #729 of 1776 rated models, from 19 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| BenchBench | 78.89 | Aggregate Score (%) | 83.8 |
| LLM2014 Logic 2024-05 | 51.54 | Score (%) | 81 |
| MMLU | 79.3 | Accuracy (%) | 79.6 |
| WildBench | 48.93 | WB Score Task-Macro | 79 |
| LLM2014 Logic 2024-06 | 51.8 | Score (%) | 75.9 |
| LLM2014 Logic 2024-07 | 51.38 | Score (%) | 61.5 |
| LLM2014 Logic 2024-08 | 47.93 | Score (%) | 54.2 |
| BigCodeBench | 37.7 | Pass@1 (%) | 47.2 |
| ZebraLogic | 18.8 | Puzzle Accuracy (%) | 42.6 |
| BenchTable | 42.8 | Total Score (%) | 41 |
| AI for Education Pedagogy - Science | 74.86 | Accuracy (%) | 38.6 |
| LLM2014 Logic 2024-09 | 45.28 | Score (%) | 37 |
Interactive version: theaggregate.ai/model?slug=yi-large · How the rankings work · Data refreshed daily, snapshot 2026-07-22.