Gemini 2.5 Flash (Non-reasoning): benchmark results

Gemini 2.5 Flash evaluated with reasoning disabled. Provider: Google. Released 2025-06-17. Access: API.

Unified ELO 1566 ± 1, rank #862 of 3078 rated models, from 293 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
TIDE Trajectory - Sudoku - AUV60.4Area Under Variation (0-100)100
EuroEval Spanish NLU - Sentiment Headlines ES52.95Sentiment classification Score (%)97.8
EuroEval Bosnian62.35Average Score (%)97.6
EuroEval Greek NLU - ScaLA EL57.72Linguistic acceptability Score (%)97.5
EuroEval Danish Knowledge - Danish Citizen Tests98.35MCC (x100)96.9
EuroEval Portuguese Common Sense Reasoning93.06Common Sense Reasoning Average Score (%)96.7
EuroEval Serbian61.96Average Score (%)96.6
EuroEval Albanian56.09Average Score (%)96.5
EuroEval Norwegian Common Sense Reasoning88.29Common Sense Reasoning Average Score (%)96.3
EuroEval Dutch73.91Average Score (%)95.9
EuroEval Icelandic NLU - Hotter and Colder Sentiment58.32Sentiment classification Score (%)95.9
EuroEval Danish Knowledge93.57Knowledge Average Score (%)95.8

Interactive version: theaggregate.ai/model?slug=gemini-2-5-flash-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-09-19.