LongDocBench - Source Relationship Recovery: leaderboard
Metric: Source relationship similarity (%; the model reads fixed TextIn parsing output and layout context of long financial reports, textbooks and papers and links each of 2,680 benchmark-localized tables and figures to its related text; normalized edit similarity to the gold text after IoU matching of objects). Source: arxiv.org. Saturation forecast: Around April 2027. 8 models tracked.
Top models
Interactive version: theaggregate.ai/benchmark?slug=longdocbench-source-relationship-recovery · How It Works · Data refreshed daily, snapshot 2026-09-26.