PASB - Short-Term Memory Extraction: leaderboard

Metric: Extraction success rate (%): share of the 40 cases in which the attacker retrieves the specified short-term context fragment (PASB direct prompt injection against the deployed OpenClaw personal agent's memory, 40 cases per task in a sandboxed testbed with canary markers; no defense; mean over repeated trials); lower is better. Source: arxiv.org. 3 models tracked.

Top models

#ModelScoreOverall rank
1Qwen 2.5 7B Instruct33.5#846
2GPT-4o Mini38.2#588
3Llama 3.1 70B Instruct41#548

Interactive version: theaggregate.ai/benchmark?slug=pasb-short-term-memory-extraction · How It Works · Data refreshed daily, snapshot 2026-10-11.