SCAM: leaderboard
1Password's open-source agent-security benchmark. Agents run 30 everyday scenarios seeded with phishing, credential-theft and social-engineering traps, three runs per phase; the score is the share of scenarios handled safely, and a companion count tracks irreversible critical failures such as submitting credentials to a phishing page.
Metric: Safe Behaviour Rate (%). Source: 1password.github.io. Status: saturation imminent. 8 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | Claude Opus 4.6 | 92 |
| 2 | GPT-5.2 | 81 |
| 3 | Gemini 3 Flash | 76 |
| 4 | Claude Haiku 4.5 | 65 |
| 5 | Claude Sonnet 4 | 49 |
| 6 | GPT-4.1 | 38 |
| 7 | GPT-4.1 Mini | 36 |
| 8 | Gemini 2.5 Flash | 35 |
Interactive version: theaggregate.ai/benchmark?slug=scam · How It Works · Data refreshed daily, snapshot 2026-09-05.