OpenCompass Language - Traditional Cultural Understanding - Chinese (CompassBench 2404): leaderboard

Metric: Score (%). Source: rank.opencompass.org.cn. 31 models tracked.

Top models

#ModelScore
1OrionStar-Yi-34B-Chat65.3
2Yi 34B (Chat)59.3
3Qwen-72B-Chat58
4Qwen-14B-Chat54
5Yi 6B (Chat)47.7
6Qwen-7B Chat42.3
7GPT-4 Turbo (Preview)36
8Baichuan2-13B-Chat32.3
9Baichuan2-7B-Chat32
10deepseek-llm-7B-chat26.7
11WizardLM-70B-V1.012.7
12WizardLM-13B-V1.27
13Llama 2 70B Chat6.3
14Mistral 7B Instruct (v0.2)5.3
15Llama 2 7B Chat4.3

Interactive version: theaggregate.ai/benchmark?slug=opencompass-language-traditional-cultural-understanding-chinese-compassbench-2404 · How It Works · Data refreshed daily, snapshot 2026-09-19.