OmniTraffic - View-BEV Mapping: leaderboard
Metric: Accuracy (%) on view to bird-eye-view mapping (map between a perspective view and the BEV layout), multiple-choice VQA on the 3,204-item human-verified OmniTraffic test set (12 reconstructed simulated intersections plus real roadside surveillance from South Korea and Tianjin), greedy decoding (at most 500 tokens), exact match against the answer letter with up to three attempts for invalid output (still invalid counts as wrong); higher is better. Source: arxiv.org. Saturation forecast: Around 2031. 11 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | Grok 4 | 39.2 |
| 2 | Gemini 3 Pro | 34 |
| 3 | GPT-4o | 32.8 |
| 4 | Claude Sonnet 4.5 (Thinking) | 29.6 |
| 5 | Qwen 3 VL 235B A22B (Thinking) | 28.4 |
| 6 | GPT-5.2 | 25.6 |
| 7 | Gemini 2.5 Pro | 9.2 |
Interactive version: theaggregate.ai/benchmark?slug=omnitraffic-view-bev-mapping · How It Works · Data refreshed daily, snapshot 2026-09-29.