GPT-5.6 Terra (Max) — benchmark results

GPT-5.6 Terra evaluated at the max reasoning-effort setting. Provider: OpenAI. Released 2026-07-09. Access: API.

Unified ELO 2051 ± 39, rank #14 of 1776 rated models, from 51 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Aikido CVE Rediscovery83.3Recall (%)100
AA CritPt30Accuracy (%)99.8
Artificial Analysis Intelligence Index54.95Intelligence Index98.9
AA Long Context Reasoning74Accuracy (%)98.3
AA Humanity's Last Exam41.8Accuracy (%)98.1
AA GPQA Diamond92.53Accuracy (%)97.6
Tau3 Banking31.75Success Rate (%)97.6
ARC-AGI-196.5Accuracy (%)97.5
AA Terminal-Bench Hard57.58Accuracy (%)97.3
OTIS Mock AIME 2024-2599.72Accuracy (%)97.1
AA SciCode53.94Accuracy (%)97
AA Omniscience - Software Engineering (SWE) - Go74Accuracy (%)96.7

Interactive version: theaggregate.ai/model?slug=gpt-5-6-terra-max · How the rankings work · Data refreshed daily, snapshot 2026-07-22.