GPT-5 Nano (2025-08-07) (Medium): benchmark results

Provider: OpenAI. Released 2025-08-07. Access: API.

Unified ELO 1550 ± 32, rank #863 of 2131 rated models, from 21 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
BeQu - Experiment 1 - Entailment Precision88Entailment Precision (%)100
MATH Level 595.24Accuracy (%)87.9
OTIS Mock AIME 2024-2574.17Accuracy (%)57.6
Epoch AI - GPQA Diamond67.42Accuracy (%)40.9
CommunityFact (Closed-Input) - Finance (English)57.77Macro-F1 (%) over True and False claims on the temporally he35.7
CommunityFact (Closed-Input) - Politics (English)55.47Macro-F1 (%) over True and False claims on the temporally he35.7
CommunityFact (Closed-Input) - Finance (French)52.56Macro-F1 (%) over True and False claims on the temporally he28.6
CommunityFact (Closed-Input) - Finance (Japanese)61.8Macro-F1 (%) over True and False claims on the temporally he28.6
CommunityFact (Closed-Input) - Finance (Portuguese)53.14Macro-F1 (%) over True and False claims on the temporally he28.6
CommunityFact (Closed-Input) - Finance (Spanish)51.29Macro-F1 (%) over True and False claims on the temporally he28.6
CommunityFact (Closed-Input) - Politics (Portuguese)55.35Macro-F1 (%) over True and False claims on the temporally he28.6
VPCT35.4Accuracy (%)25.6

Interactive version: theaggregate.ai/model?slug=gpt-5-nano-2025-08-07-medium · How It Works · Data refreshed daily, snapshot 2026-10-09.