Gemini 3 Pro (Low): benchmark results

Provider: Google. Released 2025-11-18. Access: API.

Unified ELO 1765 ± 23, rank #201 of 2133 rated models, from 24 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
BaFCo - Coarse Layout Analysis (CoT)25.13Mean average precision at IoU 0.3 (0-100): class-wise averag100
GISA - List EM27.08Exact match (%) on the 48 list-type queries (an ordered list96.7
GISA - Set F163.82F1 (%) between the predicted and gold sets on the 50 set-typ93.3
Korean CSAT 2026 (Easy Mode) - English100Points (out of 100)90.5
VitaBench30Cross-Scenario Avg@4 (%)90
MCPMark50.79Pass@1 (%)89.5
BaFCo - Coarse Layout Analysis (Zero-shot)25.78Mean average precision at IoU 0.3 (0-100): class-wise averag88.9
BaFCo - Layout Analysis (CoT)8.53Mean average precision at IoU 0.3 (0-100): class-wise averag88.9
BaFCo - Layout Analysis (Zero-shot)9Mean average precision at IoU 0.3 (0-100): class-wise averag88.9
Korean CSAT 2026 (Easy Mode) - Chemistry I48Points (out of 50)80.1
GISA - Table Item F164.93Item-level F1 (%) between predicted and gold table cells on 80
Korean CSAT 2026 (Easy Mode) - Physics I42Points (out of 50)78.7

Interactive version: theaggregate.ai/model?slug=gemini-3-pro-low · How It Works · Data refreshed daily, snapshot 2026-10-11.