ContextEcho — leaderboard

Metric: Delta (self-reported). Source: benchmarklist.com. 23 models tracked.

Top models

#ModelScore
1Claude Haiku 4.583
2Claude Sonnet 4.672
3GPT-5 Mini65
4DeepSeek V365
5Claude Sonnet 4.563
6Mistral Large50
7GPT-4.147
8Command A47
9Gemini 2.5 Flash43
10Qwen 3 235B A22B40
11Gemini 2.5 Pro38
12GPT-4o38
13GPT-528
14Mistral Small27
15Llama 3.3 70B Instruct-3

Interactive version: theaggregate.ai/benchmark?slug=contextecho · How the rankings work · Data refreshed daily, snapshot 2026-07-22.