GPT-5.2 (Low) — benchmark results

GPT-5.2 evaluated at the low reasoning-effort setting. Provider: OpenAI. Released 2025-12-11. Access: API.

Unified ELO 1755 ± 27, rank #187 of 1776 rated models, from 41 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
UGI - Natural Intelligence55.79NatInt Score92.8
LLM Chess (Saplin)782.5ELO90
UGI - Writing43.85Writing Score84.5
Epoch AI - ECI153.72ECI Score82
Pencil Puzzle Bench - Nurimisaki13.3Direct-ask Success Rate (%)81
NonoBench46.7Overall Accuracy (%)79.8
Pencil Puzzle Bench - Slitherlink6.7Direct-ask Success Rate (%)74
Pencil Puzzle Bench - Light Up6.7Direct-ask Success Rate (%)71
Pencil Puzzle Bench - Shikaku6.7Direct-ask Success Rate (%)71
FrontierMath - Tiers 1-326.55Accuracy (%, 290 problems)69.7
OTIS Mock AIME 2024-2578.89Accuracy (%)62.8
Vectara Hallucination Leaderboard91.6Factual Consistency Rate (%)60.6

Interactive version: theaggregate.ai/model?slug=gpt-5-2-low · How the rankings work · Data refreshed daily, snapshot 2026-07-22.