UniRef-UAV (Text+Image): leaderboard

Metric: Precision@(F1=1, IoU>=0.5) (%; share of the 10,322 text+image test queries whose predicted box set exactly matches the referred target set under one-to-one IoU>=0.5 matching, an empty prediction being correct only for no-target queries; zero-shot, model-specific grounding prompt, unparseable output read as an empty prediction). Source: arxiv.org. Saturation forecast: Around December 2026. 13 models tracked.

Top models

#ModelScore
1Qwen 3.6 27B58.5
2Gemma 4 31B (IT)46.2

Interactive version: theaggregate.ai/benchmark?slug=uniref-uav-text-plusimage · How It Works · Data refreshed daily, snapshot 2026-09-29.