Gemini 2.5 Pro — benchmark results
Google Gemini 2.5 Pro model for advanced reasoning and multimodal tasks. Provider: Google. Released 2025-03-25. Access: API.
Unified ELO 1749 ± 6, rank #198 of 1776 rated models, from 1794 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AI for Education Visual Maths - Algebra | 100 | Accuracy (%) | 100 |
| Active Evidence-Seeking and Diagnostic Reasoni | 49.4 | Task 2: Active Seeking Exact Accuracy (self-reported) | 100 |
| BIRD-Interact (c-Interact) | 20.92 | Normalized Reward | 100 |
| Chartographer | 94.7 | ChartQA OA (self-reported) | 100 |
| DeepResearch Bench - Effective Citations | 165.34 | Effective Citations | 100 |
| Diplomacy: Betrayal Tendency | 100 | Betrayal Rate (%) | 100 |
| Episodic Memory | 96.8 | Simple Recall (%) | 100 |
| EuroEval Czech | 70.02 | Average Score (%) | 100 |
| EuroEval Estonian | 62.38 | Average Score (%) | 100 |
| EuroEval Hungarian | 67.51 | Average Score (%) | 100 |
| EuroEval Norwegian Knowledge | 87.61 | Knowledge Average Score (%) | 100 |
| Evals for Every Language - Language bho | 68.98 | Average Score (%) | 100 |
Interactive version: theaggregate.ai/model?slug=gemini-2-5-pro · How the rankings work · Data refreshed daily, snapshot 2026-07-22.