PatRe - Office Action Decision (OA-DP): leaderboard
Metric: Decision accuracy (%): the generated office action's decision (allowance, non-final or final rejection, Ex parte Quayle) matches the examiner's, office action generation under direct prompting: the model writes the office action from the claims (and any preceding rebuttal) with no prior art, over PatRe (480 recent USPTO patent examination histories across all eight IPC sections, with office actions, rebuttals, claim versions and cited references), temperature 0; higher is better. Source: arxiv.org. Saturation forecast: Around August 2027. 10 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | GPT-5 Mini | 51.4 |
| 2 | Gemini 2.5 Flash | 50 |
| 3 | Qwen 3.5 27B | 48.8 |
| 4 | DeepSeek V3.2 | 47.6 |
| 5 | Gemma 3 12B (IT) | 45.1 |
| 6 | Qwen 3.5 9B | 41.7 |
| 7 | Llama 3.1 8B Instruct | 41.4 |
| 8 | Gemma 3 27B (IT) | 39.5 |
| 9 | GPT-4o Mini | 24.4 |
| 10 | Llama 3.3 70B Instruct | 10.2 |
Interactive version: theaggregate.ai/benchmark?slug=patre-office-action-decision-oa-dp · How It Works · Data refreshed daily, snapshot 2026-10-07.