OmniHandwritingOCR - Chinese Text: leaderboard

Metric: 1-NED (%; one minus the normalized Levenshtein distance between the transcription and the ground truth, averaged over 12,529 Chinese handwritten text images (CASIA-HWDB and newly collected student writing); zero-shot Markdown transcription from one image with one task-neutral instruction for every system; task-specific normalization and tokenization, no semantic correction). Source: arxiv.org. Saturation forecast: Estimated already saturated. 13 models tracked.

Top models

#ModelScore
1InternVL3-78B79.91
2GPT-4o60.66
3Gemma 3 27B (IT)29.28

Interactive version: theaggregate.ai/benchmark?slug=omnihandwritingocr-chinese-text · How It Works · Data refreshed daily, snapshot 2026-09-29.