SCAM - With Security Skill: leaderboard

The same 30 SCAM scenarios re-run with 1Password's plain-text security skill prepended to the agent's system prompt, which teaches it to verify domains, inspect content and protect credentials before acting. Read against the base SCAM board it measures how much of an agent's security gap is closed by instructions alone.

Metric: Safe Behaviour Rate (%). Source: 1password.github.io. Status: saturated. 8 models tracked.

Top models

#ModelScore
1Gemini 3 Flash99
2Claude Opus 4.698
3Claude Sonnet 498
4Claude Haiku 4.598
5GPT-5.297
6GPT-4.196
7GPT-4.1 Mini95
8Gemini 2.5 Flash95

Interactive version: theaggregate.ai/benchmark?slug=scam-with-security-skill · How It Works · Data refreshed daily, snapshot 2026-09-05.