Yi Large — benchmark results

01.AI's closed API flagship chat model with strong multilingual performance, the Chinese startup's top proprietary offering of 2024. Provider: 01.AI. Released 2024-06-25. Access: API.

Unified ELO 1519 ± 15, rank #729 of 1776 rated models, from 19 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
BenchBench78.89Aggregate Score (%)83.8
LLM2014 Logic 2024-0551.54Score (%)81
MMLU79.3Accuracy (%)79.6
WildBench48.93WB Score Task-Macro79
LLM2014 Logic 2024-0651.8Score (%)75.9
LLM2014 Logic 2024-0751.38Score (%)61.5
LLM2014 Logic 2024-0847.93Score (%)54.2
BigCodeBench37.7Pass@1 (%)47.2
ZebraLogic18.8Puzzle Accuracy (%)42.6
BenchTable42.8Total Score (%)41
AI for Education Pedagogy - Science74.86Accuracy (%)38.6
LLM2014 Logic 2024-0945.28Score (%)37

Interactive version: theaggregate.ai/model?slug=yi-large · How the rankings work · Data refreshed daily, snapshot 2026-07-22.