GPT-5.4 Nano (Low): benchmark results

Provider: OpenAI. Released 2026-03-17. Access: API.

Unified ELO 1539 ± 30, rank #905 of 2131 rated models, from 15 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
FrontierMath - Tiers 1-3 (v2)20.35Accuracy (%, 285 private v2 problems)31.3
MageBench S21584Combined 1v1 Elo20
ARC-AGI-21.53Accuracy (%)18.2
ARC-AGI-118.33Accuracy (%)15
ObviousBench52.78Answer pass³ (%)13.2
T1-Bench - Tool Call F173.66Tool-name F1 (%) of the assistant's tool calls against the g9.1
Context Arena21.87Average Score (%)7.3
o11y-bench - Pass@350.79Tasks passed on at least one of three attempts, Pass@3 (%)3.5
o11y-bench - Pass^319.05Tasks passed on all three attempts, Pass^3 (%)1.8
Equation-Suffix Prediction (Kimi K2.6 Scorer)0.05Likelihood lift (mean clipLL2 per target token over the same0
Equation-Suffix Prediction (Qwen3-8B Scorer)0.08Likelihood lift (mean clipLL2 per target token over the same0
T1-Bench20.38Pass@3 (%): share of conversations solved in at least one of0

Interactive version: theaggregate.ai/model?slug=gpt-5-4-nano-low · How It Works · Data refreshed daily, snapshot 2026-10-09.