GPT-5.4 Mini — benchmark results
OpenAI's smaller, faster GPT-5.4 tier for cost-efficient reasoning workloads. Provider: OpenAI. Released 2026-03-17. Access: API.
Unified ELO 1696 ± 12, rank #276 of 1776 rated models, from 256 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AGC-Bench - hypobench | 0.48 | Dataset z-score | 100 |
| CalBench | 149 | Excess (self-reported) | 100 |
| GACL - Battleship | 85.94 | Normalized Score (0-100) | 100 |
| TempGlitch | 52.4 | Acc. (1 FPS) (self-reported) | 100 |
| AGC-Bench - pun_eval | 1.4 | Dataset z-score | 97.6 |
| AGC-Bench - humor_transfer | 0.97 | Dataset z-score | 96.3 |
| LLM-as-a-Reviewer | 99.7 | Low (NeurIPS 2022) (self-reported) | 95.5 |
| AGC-Bench - pollux_creativity | 0.92 | Dataset z-score | 95.1 |
| CritPt | 10 | Accuracy (self-reported) | 94.5 |
| SLR-Bench - Hard | 28 | Accuracy (%) | 94.5 |
| AA-LCR | 69.3 | Score (self-reported) | 93.5 |
| Chartographer | 94 | ChartQA OA (self-reported) | 92.9 |
Interactive version: theaggregate.ai/model?slug=gpt-5-4-mini · How the rankings work · Data refreshed daily, snapshot 2026-07-22.