GPT-5.4 Nano: benchmark results
OpenAI's smallest, fastest GPT-5.4 tier. Provider: OpenAI. Released 2026-03-17. Access: API.
Unified ELO 1599 ± 1, rank #270 of 1392 rated models, from 171 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Honeyval | 96.9 | Interaction Length (self-reported) | 100 |
| MLX Benchmark V2 - MLX Embeddings LoRA | 86.67 | Accuracy (%) | 100 |
| Vectara Hallucination Leaderboard | 96.9 | Factual Consistency Rate (%) | 99 |
| AGC-Bench - humor_transfer | 0.97 | Dataset z-score | 96.3 |
| AGC-Bench - hypobench | 0.42 | Dataset z-score | 95.1 |
| MLX Benchmark V2 - Coding | 75.76 | Accuracy (%) | 95 |
| MLX Benchmark V2 - Conceptual | 92.31 | Accuracy (%) | 90 |
| MLX Benchmark V2 - Very Hard | 82 | Accuracy (%) | 90 |
| RAI-Bench - RAG Robustness (LC Abstention) | 92 | Rate (%) | 89.9 |
| AGC-Bench - pun_eval | 0.87 | Dataset z-score | 87.8 |
| AGC-Bench - thenextchapter | 0.93 | Dataset z-score | 86.6 |
| AGC-Bench - pron_vs_prompt | 1 | Dataset z-score | 86.4 |
Interactive version: theaggregate.ai/model?slug=gpt-5-4-nano · How It Works · Data refreshed daily, snapshot 2026-09-05.