UniRef-UAV (Text+Image): leaderboard
Metric: Precision@(F1=1, IoU>=0.5) (%; share of the 10,322 text+image test queries whose predicted box set exactly matches the referred target set under one-to-one IoU>=0.5 matching, an empty prediction being correct only for no-target queries; zero-shot, model-specific grounding prompt, unparseable output read as an empty prediction). Source: arxiv.org. Saturation forecast: Around December 2026. 13 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | Qwen 3.6 27B | 58.5 |
| 2 | Gemma 4 31B (IT) | 46.2 |
Interactive version: theaggregate.ai/benchmark?slug=uniref-uav-text-plusimage · How It Works · Data refreshed daily, snapshot 2026-09-29.