PaveVQA - Quantification MAPE: leaderboard
Metric: Mean absolute percentage error (%, no upper limit) of quantitative answers (distress area, length and severity measures), of the zero-shot model on PaveVQA, the vision-language question answering part of PaveBench (real highway pavement images, single-turn, multi-turn and expert-corrected questions); lower is better. Source: arxiv.org. Saturation forecast: Rough model projection: around 2026. 3 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | Qwen2.5-VL-3B | 116.2 |
| 2 | LLaVA-OneVision-7B | 158.24 |
| 3 | DeepSeek-VL2-small | 531.21 |
Interactive version: theaggregate.ai/benchmark?slug=pavevqa-quantification-mape · How It Works · Data refreshed daily, snapshot 2026-10-07.