GPT-5.2 (2025-12-11) (Medium): benchmark results

Provider: OpenAI. Released 2025-12-11. Access: API.

Unified ELO 1738 ± 26, rank #259 of 2131 rated models, from 23 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Epoch AI - Algotune2.05Score100
SWE-rebench59.25Resolved (%)90.6
Chess Puzzles (Epoch AI)40Accuracy (%)88.5
PLCC - Vocabulary86Accuracy (%)86.2
OTIS Mock AIME 2024-2593.89Accuracy (%)83
PLCC - Geography94Accuracy (%)82.6
Terminal-Bench 2.064.9Accuracy (%)82.1
Context-Bench Skills77.55Task Completion (%)81
Epoch AI - GPQA Diamond87.88Accuracy (%)79.9
PLCC - Overall85Mean category accuracy (%)79.9
PLCC - History90Accuracy (%)78.8
PLCC - Grammar82Accuracy (%)77.5

Interactive version: theaggregate.ai/model?slug=gpt-5-2-2025-12-11-medium · How It Works · Data refreshed daily, snapshot 2026-10-09.