Yi Large (Preview) — benchmark results
01.AI preview Yi Large model row. Provider: 01.AI. Released 2024-06-17. Access: API.
Unified ELO 1504 ± 49, rank #785 of 1776 rated models, from 11 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| BenchBench | 87.14 | Aggregate Score (%) | 94.1 |
| WildBench | 55.29 | WB Score Task-Macro | 93.5 |
| AlpacaEval 2.0 | 51.89 | LC Win Rate (%) | 88.3 |
| MixEval | 56.8 | Score | 80.4 |
| HELM NaturalQuestions (Closed) | 42.77 | F1 (%) | 72.2 |
| HELM Lite | 51.21 | Mean win rate (self-reported) | 53.2 |
| HELM (Stanford) | 47.1 | Mean Win Rate (%) | 45.6 |
| ZebraLogic | 18.9 | Puzzle Accuracy (%) | 44.3 |
| HELM WMT 2014 | 17.61 | BLEU-4 (%) | 34.4 |
| HELM NaturalQuestions (Open) | 58.64 | F1 (%) | 8.9 |
| HELM NarrativeQA | 37.28 | F1 (%) | 3.3 |
Interactive version: theaggregate.ai/model?slug=yi-large-preview · How the rankings work · Data refreshed daily, snapshot 2026-07-22.