GPT-5.4 (2026-03-05) (Low): benchmark results

Provider: OpenAI. Released 2026-03-05. Access: API.

Unified ELO 1729 ± 28, rank #277 of 2131 rated models, from 19 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
PLCC - Geography97Accuracy (%)92.2
PLCC - Art & Entertainment87Accuracy (%)90.3
PLCC - Culture & Tradition93Accuracy (%)90.3
PLCC - Overall90.5Mean category accuracy (%)90
PLCC - History93Accuracy (%)87.7
EgoGapBench58.3Accuracy (%; zero-shot multiple choice on 1,000 single-image87.5
EgoGapBench - Direct Actions40.9Accuracy (%; zero-shot multiple choice on 650 direct-action 87.5
EgoGapBench - Four Options57.3Accuracy (%; zero-shot multiple choice on 382 four-option si87.5
EgoGapBench - Indirect Actions90.6Accuracy (%; zero-shot multiple choice on 350 indirect-actio87.5
EgoGapBench - Three Options58.9Accuracy (%; zero-shot multiple choice on 618 three-option s87.5
PLCC - Grammar88Accuracy (%)86.9
PLCC - Vocabulary85Accuracy (%)84.3

Interactive version: theaggregate.ai/model?slug=gpt-5-4-2026-03-05-low · How It Works · Data refreshed daily, snapshot 2026-10-09.