Yi-34B-200K-AEZAKMI-v2: benchmark results

Provider: 01.AI. Access: Open.

Unified ELO 1524 ± 20, rank #1017 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - MMLU75.22Accuracy (%) (5-shot)96.6
Open LLM Leaderboard - MMLU-Pro45.13Score84
Open LLM Leaderboard - GPQA33.22Score81.4
Open LLM Leaderboard v1 - WinoGrande81.61Accuracy (%) (5-shot)80.1
Open LLM Leaderboard v1 - ARC Challenge67.92Normalized accuracy (%) (25-shot)78.6
Open LLM Leaderboard v1 - HellaSwag85.61Normalized accuracy (%) (10-shot)75.6
Open LLM Leaderboard v1 - GSM8K58.91Accuracy (%) (5-shot)70.7
Open LLM Leaderboard v1 - TruthfulQA MC256.74MC2 (%) (0-shot)68.6
Open LLM Leaderboard - BBH53.84Score68
Open LLM Leaderboard - IFEval45.55Score50.7
Open LLM Leaderboard - MuSR38.86Score34.5
Open LLM Leaderboard - MATH Level 55.66Score31.3

Interactive version: theaggregate.ai/model?slug=yi-34b-200k-aezakmi-v2 · How It Works · Data refreshed daily, snapshot 2026-09-23.