SMILE-Temporal: leaderboard

Metric: Laughter localization precision at IoU 0.5 (times 100, so 0-100): share of positive test segments whose predicted laughter interval overlaps a ground-truth laughter event with IoU of at least 0.5 (maximum over the segment's events); the model is prompted, zero-shot, to say whether laughter occurs in a 5-second video segment and give its start and end time, on the SMILE-Temporal test split (clips of the SMILE laughter-reasoning video dataset); higher is better. Source: arxiv.org. Saturation forecast: Around January 2027. 3 models tracked.

Top models

#ModelScore
1Gemini 3 Flash57.9
2Qwen2.5-Omni-7B31.5

Interactive version: theaggregate.ai/benchmark?slug=smile-temporal · How It Works · Data refreshed daily, snapshot 2026-10-07.