Seed 2.0 Pro — benchmark results
ByteDance's flagship Seed 2.0 Pro reasoning and agent model. Provider: ByteDance. Released 2026-02-16. Access: API.
Unified ELO 1687 ± 24, rank #299 of 1776 rated models, from 36 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| VitaBench 2.0 | 47.4 | Avg@4 Full Context (self-reported) | 94.7 |
| MMGist | 58.2 | Macro â (self-reported) | 92.3 |
| Position Bias (Lechmazur) | 28 | Order Flip % (lower is better) | 91.4 |
| ZeroEval GPQA Diamond | 88.9 | GPQA Diamond Score | 88.9 |
| CC-OCR V2 | 72.15 | Average (self-reported) | 85.7 |
| Persuasion (Lechmazur) | 1.64 | Average Persuasion Strength | 85.7 |
| ClawProBench | 61.07 | Final Score (self-reported) | 83.9 |
| OpenClawProBench | 68.3 | Overall Score (%) | 78.8 |
| VisualNeedle | 36.8 | w/ Tools Acc. (%) (self-reported) | 75 |
| LiveSecBench | 63.04 | Overall Score (%) | 71.4 |
| LLM Stats (AIME 2026) | 94.2 | Score (%) | 68.8 |
| Visual Aesthetic Benchmark (VAB) | 51.3 | Overall Top-1 ap@1 (self-reported) | 66.7 |
Interactive version: theaggregate.ai/model?slug=seed-2-0-pro · How the rankings work · Data refreshed daily, snapshot 2026-07-22.