Gemini 2.5 Flash (Preview) (Reasoning) — benchmark results
Provider: Google. Released 2025-06-17. Access: API.
Unified ELO 1721 ± 62, rank #240 of 1841 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA MATH-500 | 98.07 | Accuracy (%) | 92.2 |
| AA MMLU-Pro | 79.98 | Accuracy (%) | 66.9 |
| AA Humanity's Last Exam | 11.56 | Accuracy (%) | 65.9 |
| AA LiveCodeBench | 50.48 | Pass@1 (%) | 56.4 |
| Epoch AI - Scicode | 35.88 | Score | 55.4 |
| Artificial Analysis Intelligence Index | 17.55 | Intelligence Index | 53.6 |
| AA GPQA Diamond | 69.8 | Accuracy (%) | 52.3 |
Interactive version: theaggregate.ai/model?slug=gemini-2-5-flash-preview-reasoning · How It Works · Data refreshed daily, snapshot 2026-07-25.