GPT-6 Astra: benchmark results
Provider: OpenAI. Released 2026-09-03. Access: API.
Unified ELO 2019 ± 5, rank #4 of 1609 rated models, from 683 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA GDP.pdf | 32.2 | All Criteria Pass Rate (%) | 100 |
| AA GPQA Diamond | 96.26 | Accuracy (%) | 100 |
| AA-Omniscience Index - Health | 27 | Omniscience Index | 100 |
| AA-Omniscience Index - Software Engineering (SWE) - Go | 92 | Omniscience Index | 100 |
| AA-Omniscience Index - Software Engineering (SWE) - TypeScript | 97.78 | Omniscience Index | 100 |
| AI for Education Pedagogy | 92.66 | Accuracy (%) | 100 |
| ARC-AGI-3 | 62.71 | Accuracy (%) | 100 |
| ActiveVision | 76.5 | Accuracy (%) | 100 |
| Agent Arena - Praise vs Complaint | 34.7 | Praise vs Complaint (%) | 100 |
| Agents on Rails | 53.3 | Successful Runs (%) | 100 |
| BALROG BabaIsAI (LLM) | 100 | Progress (%) | 100 |
| BALROG Crafter (LLM) | 76.8 | Progress (%) | 100 |
Interactive version: theaggregate.ai/model?slug=gpt-6-astra · How It Works · Data refreshed daily, snapshot 2026-10-02.