Gemini 1.5 Pro (Sept): benchmark results

September 2024 Gemini 1.5 Pro snapshot, tracked when sources report the dated API model. Provider: Google. Released 2024-09-24. Access: API.

Unified ELO 1546 ± 1, rank #456 of 1392 rated models, from 18 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
RewardBench86.78Score (%)86.1
Deception Effectiveness (Lechmazur)0.94Deception Score76.5
RewardBench Chat Hard76.97Accuracy (%)76.1
LLM Public Goods Game24.64Avg. Contribution (%)73.7
RewardBench Reasoning90.22Accuracy (%)73.4
RewardBench Safety85.81Accuracy (%)66.7
Confabulation Leaderboard (Lechmazur)16.83Confabulation rate % (lower is better)61.9
RewardBench Chat94.13Accuracy (%)48.9
Deception Resistance (Lechmazur)0.62Vulnerability Score (lower is better)47.1
NYT Connections Original22.7Score (%)43.3
AA GPQA Diamond58.89Accuracy (%)33.6
Artificial Analysis Intelligence Index4.27Intelligence Index32.8

Interactive version: theaggregate.ai/model?slug=gemini-1-5-pro-sept · How It Works · Data refreshed daily, snapshot 2026-09-05.