Hear2Act (Spoken, Prosody-Mediated Feedback): leaderboard

Metric: Optimal-solution rate (%; share of dialogues ending with the first-tier option, using the final recommendation if the turn budget runs out; 480 persona-grounded consumer-service scenarios with a hidden user concern and verifiable outcomes; prosody-mediated feedback: the concern is carried by the delivery of Qwen3-TTS speech with unchanged words; the spoken assistant receives the user audio and the transcript; one rollout per scenario for Qwen2.5-Omni-7B, three for Qwen2-Audio-7B-Instruct). Source: arxiv.org. Saturation forecast: Around July 2027. 2 models tracked.

Top models

#ModelScore
1Qwen2.5-Omni-7B15.9
2Qwen2-Audio-7B-Instruct14.7

Interactive version: theaggregate.ai/benchmark?slug=hear2act-spoken-prosody-mediated-feedback · How It Works · Data refreshed daily, snapshot 2026-09-29.