fr-bench-pdf2md - Forms: leaderboard
Metric: Unit-test pass rate (%) of the model's Markdown conversion on fr-bench-pdf2md's forms (semi-structured forms with handwritten entries written by participants) pages (difficult French PDF pages from CCPDF and Gallica, chosen where two OCR models disagreed most), checking text presence, reading order and table structure after category-specific normalization; higher is better. Source: arxiv.org. Saturation forecast: Around January 2027. 17 models tracked.
Top models
| # | Model | Score | Overall rank |
|---|---|---|---|
| 1 | Gemini 3 Pro (Preview) | 72.5 | #64 |
| 2 | Gemini 3 Flash (Preview) | 68.4 | #78 |
| 3 | GPT-5.2 | 48.1 | #105 |
| 4 | GPT-5 Mini | 41.6 | #176 |
| 5 | Gemini 2.5 Flash Lite | 38.8 | #413 |
No result here: #3 Claude Opus 5.5, #5 GPT-6 Astra, #8 Claude Fable 5.1.
Interactive version: theaggregate.ai/benchmark?slug=fr-bench-pdf2md-forms · How It Works · Data refreshed daily, snapshot 2026-10-11.