AA-Omniscience Net Score — leaderboard

AA-Omniscience factuality results reported by Anthropic as net score: correct responses minus incorrect responses, with abstentions scoring zero.

Metric: Net score (self-reported). Source: benchmarklist.com. Status: saturation imminent. 7 models tracked.

Top models

#ModelScore
1Claude Mythos 553
2Claude Mythos Preview50
3Claude Opus 4.837
4Claude Opus 4.735
5Claude Opus 4.621
6Claude Opus 4.515
7Claude Sonnet 4.614

Interactive version: theaggregate.ai/benchmark?slug=aa-omniscience-net-score · How the rankings work · Data refreshed daily, snapshot 2026-07-22.