CCLUE — leaderboard
Classical Chinese Language Understanding Evaluation benchmark covering segmentation and punctuation, named entity recognition, classification, sentiment, and classical-modern retrieval.
Metric: Average Score (%). Source: ethan-yt.github.io. 3 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | guwenbert-base | 77.55 |
| 2 | guwenbert-base-fs | 77.21 |
| 3 | chinese-roberta-wwm-ext | 70.88 |
Interactive version: theaggregate.ai/benchmark?slug=cclue · How the rankings work · Data refreshed daily, snapshot 2026-07-22.