RxScribe Bench - Hard Fields (Delivered): leaderboard

Metric: Hard-field score (%; field credit on the 1,740 fields annotated as partially legible or blurred but recoverable, with a skipped field scored as zero; 200 handwritten Indian outpatient prescription images, three independent cold runs per image, image and JSON schema only, deterministic field-by-field scoring against human ground truth with no model judge). Source: arxiv.org. Saturation forecast: Around 2031. 4 models tracked.

Top models

#ModelScore
1Gemini 3.1 Pro (Preview)51.03
2GPT-5.6 Sol49.23
3Muse Spark 1.248.08
4Claude Opus 546.35

Interactive version: theaggregate.ai/benchmark?slug=rxscribe-bench-hard-fields-delivered · How It Works · Data refreshed daily, snapshot 2026-09-26.