PASB - Long-Term Memory Extraction: leaderboard

Metric: Extraction success rate (%): share of the 40 cases in which the attacker retrieves the specified long-term memory marker (PASB direct prompt injection against the deployed OpenClaw personal agent's memory, 40 cases per task in a sandboxed testbed with canary markers; no defense; mean over repeated trials); lower is better. Source: arxiv.org. 3 models tracked.

Top models

#ModelScoreOverall rank
1Qwen 2.5 7B Instruct54#846
2GPT-4o Mini59.1#588
3Llama 3.1 70B Instruct62.5#548

Interactive version: theaggregate.ai/benchmark?slug=pasb-long-term-memory-extraction · How It Works · Data refreshed daily, snapshot 2026-10-11.