AgentSocialBench - Mediated Communication Leakage: leaderboard

Metric: Privacy leakage rate (0-1, share of designated private items leaked partially or fully) of LLM agents in mediated communication (MC) scenarios, where an agent brokers a conversation between its user and another person, under the L0 (unconstrained) privacy instruction level, judged by Claude Opus 4.6; lower is better. Source: arxiv.org. Saturation forecast: Around January 2027. 8 models tracked.

Top models

#ModelScore
1Claude Sonnet 4.60.21
2DeepSeek V3.20.21
3GPT-5 Mini0.23
4Claude Sonnet 4.50.24
5MiniMax-M2.10.25
6Qwen 3 235B A22B0.26
7Claude Haiku 4.50.27
8Kimi K2.50.3

Interactive version: theaggregate.ai/benchmark?slug=agentsocialbench-mediated-communication-leakage · How It Works · Data refreshed daily, snapshot 2026-10-07.