GPT-5.4 Mini (2026-03-17): benchmark results

Provider: OpenAI. Released 2026-03-17. Access: API.

Unified ELO 1721 ± 10, rank #155 of 1639 rated models, from 613 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
AuthMem-Bench (Task Success Rate)56.6Task success rate (%; share of 350 authorized counterpart ca100
EuroEval Bulgarian NLU73.19NLU Average Score (%)100
EuroEval Latvian NLU - Latvian Twitter Sentiment56.55Sentiment classification Score (%)100
EuroEval Polish NLU - KPWr NER78.64Named entity recognition Score (%)100
EuroEval Slovene NLU - SSJ500K NER83.51Named entity recognition Score (%)100
EuroEval Swedish Knowledge88.01Knowledge Average Score (%)100
EuroEval Swedish NLU - SUC383.44Named entity recognition Score (%)100
EuroEval Ukrainian NLU - NER UK83.05Named entity recognition Score (%)100
EuroEval Portuguese NLU - ScaLA PT74.28Linguistic acceptability Score (%)99.8
EuroEval German NLU - ScaLA DE72.72Linguistic acceptability Score (%)99.7
EuroEval Czech NLU62.29NLU Average Score (%)99.6
EuroEval Albanian NLU - MMS SQ30.83Sentiment classification Score (%)99.5

Interactive version: theaggregate.ai/model?slug=gpt-5-4-mini-2026-03-17 · How It Works · Data refreshed daily, snapshot 2026-10-09.