GPT-5 Mini (2025-08-07) (Medium): benchmark results

Provider: OpenAI. Released 2025-08-07. Access: API.

Unified ELO 1681 ± 16, rank #412 of 2131 rated models, from 47 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Swallow - English MT-Bench - Average86.9Judge Score (normalized, %)100
Swallow - Japanese MT-Bench - Coding87.9Judge Score (normalized, %)100
Swallow - Japanese MT-Bench - STEM83.8Judge Score (normalized, %)100
Swallow - English MT-Bench - Coding86.1Judge Score (normalized, %)98.5
Swallow - English MT-Bench - Extraction84.5Judge Score (normalized, %)98.5
Swallow - English MT-Bench - Humanities83.4Judge Score (normalized, %)98.5
Swallow - English MT-Bench - Reasoning91Judge Score (normalized, %)98.5
Swallow - English MT-Bench - STEM84.9Judge Score (normalized, %)98.5
Swallow - English MT-Bench - Writing80.5Judge Score (normalized, %)98.5
Swallow - English MT-Bench - Roleplay85.3Judge Score (normalized, %)97
Swallow - Japanese MT-Bench - Average83Judge Score (normalized, %)97
Swallow - Japanese MT-Bench - Humanities80.3Judge Score (normalized, %)97

Interactive version: theaggregate.ai/model?slug=gpt-5-mini-2025-08-07-medium · How It Works · Data refreshed daily, snapshot 2026-10-09.