SCAM - With Security Skill: leaderboard
The same 30 SCAM scenarios re-run with 1Password's plain-text security skill prepended to the agent's system prompt, which teaches it to verify domains, inspect content and protect credentials before acting. Read against the base SCAM board it measures how much of an agent's security gap is closed by instructions alone.
Metric: Safe Behaviour Rate (%). Source: 1password.github.io. Status: saturated. 8 models tracked.
Top models
| # | Model | Score |
|---|---|---|
| 1 | Gemini 3 Flash | 99 |
| 2 | Claude Opus 4.6 | 98 |
| 3 | Claude Sonnet 4 | 98 |
| 4 | Claude Haiku 4.5 | 98 |
| 5 | GPT-5.2 | 97 |
| 6 | GPT-4.1 | 96 |
| 7 | GPT-4.1 Mini | 95 |
| 8 | Gemini 2.5 Flash | 95 |
Interactive version: theaggregate.ai/benchmark?slug=scam-with-security-skill · How It Works · Data refreshed daily, snapshot 2026-09-05.