OSCBench - Regular Scenarios: leaderboard

Metric: Object state change score (mean of accuracy and consistency) on OSCBench's 108 regular scenarios (common cooking action-object pairs), human ratings (three raters, 1-5 Likert, printed on a 0-1 scale, x100), one generated video per scenario; higher is better. Source: arxiv.org. Saturation forecast: Rough model projection: around 2026. 6 models tracked.

Top models

#ModelScoreOverall rank
1Veo-3.1-Fast79.7
2Kling 2.5 Turbo74.4
3Wan2.263.5
4HunyuanVideo-1.557.2
5HunyuanVideo47.2
6Open-Sora-2.041

Interactive version: theaggregate.ai/benchmark?slug=oscbench-regular-scenarios · How It Works · Data refreshed daily, snapshot 2026-10-11.