OmniStarPro-RNG (Online): leaderboard
Metric: Semantic correctness (0-10; GPT-4o judge score of the earliest response in each ground-truth clip, mean of semantic accuracy, language quality and information completeness; the model decides by itself when to speak, 1,000 evaluation streams of the OmniStarPro-Live partition). Source: arxiv.org. Saturation forecast: Rough model projection: around 2026. 5 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | LiveStarPro | 3.27 |
| 2 | LiveStar | 3.19 |
| 3 | MMDuet | 1.93 |
| 4 | VideoLLM-online | 1.68 |
| 5 | VideoLLM-MoD | 1.66 |
Interactive version: theaggregate.ai/benchmark?slug=omnistarpro-rng-online · How It Works · Data refreshed daily, snapshot 2026-09-26.