Yi-1.5-34B-Chat-16K: benchmark results

16K-context variant of 01.AI's Yi-1.5-34B-Chat, extending the standard 4K chat model for longer documents and conversations (May 2024). Provider: 01.AI. Released 2024-05-15. Access: Open.

Unified ELO 1485 ± 1, rank #760 of 1392 rated models, from 41 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open Chinese LLM - ARC Challenge63.74Accuracy (%)93.9
Open Chinese LLM Leaderboard69.14Average Score (%)93.5
Open Chinese LLM - C-Eval Semantic90.09Accuracy (%)93.2
Open Chinese LLM - HellaSwag69.35Accuracy (%)91.4
Open Chinese LLM - GSM8K66.72Accuracy (%)90.8
Open Chinese LLM - WinoGrande69.53Accuracy (%)90.5
Open Chinese LLM - CMMLU70.03Accuracy (%)84.9
Open LLM Leaderboard - BBH44.54Score84.8
Open LLM Leaderboard - MMLU-Pro39.38Score84.5
Open LLM Leaderboard - GPQA11.74Score83.7
MERA - RCB57.76Accuracy (%)80.4
Open LLM Leaderboard - MuSR13.74Score75.6

Interactive version: theaggregate.ai/model?slug=yi-1-5-34b-chat-16k · How It Works · Data refreshed daily, snapshot 2026-09-05.