Yi 6B (Base): benchmark results

Provider: 01.AI. Released 2023-11-02. Access: Open.

Unified ELO 1334 ± 1, rank #1296 of 1392 rated models, from 84 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open Chinese LLM - WinoGrande64.56Accuracy (%)67.4
Open Chinese LLM - C-Eval Semantic74.77Accuracy (%)64.7
Open Chinese LLM - CMMLU56.54Accuracy (%)61.7
URIAL-Bench - STEM7.7Judge Score (0-10)61.1
URIAL-Bench - Humanities8.78Judge Score (0-10)55.6
EuroEval Portuguese NLU - MultiWikiQA PT69.73Reading comprehension Score (%)55.1
URIAL-Bench - Roleplay6.95Judge Score (0-10)50
MERA - ruHumanEval9.09pass@1 (%)45.4
URIAL-Bench - Reasoning3.5Judge Score (0-10)44.4
WinoGrande71.3Accuracy (%)42.5
Open Chinese LLM - HellaSwag59.03Accuracy (%)42.1
MERA - ruCodeEval5.49pass@1 (%)41.6

Interactive version: theaggregate.ai/model?slug=yi-6b-base · How It Works · Data refreshed daily, snapshot 2026-09-05.