GPT-4 (0613) — benchmark results
June 13, 2023 GPT-4 API snapshot, kept separate from other GPT-4 rows when sources report it explicitly. Provider: OpenAI. Released 2023-06-13. Access: API.
Unified ELO 1565 ± 15, rank #578 of 1776 rated models, from 75 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| FastEval | 77.78 | Total Score | 100 |
| HELM v2 Lite - HumanEval (Code) | 73 | Pass@1 (%) | 100 |
| HELM v2 Lite - LegalBench | 71.01 | Exact Match (%) | 100 |
| HELM v2 Lite - MedQA | 88.46 | Exact Match (%) | 100 |
| InfiBench | 70.64 | Score (%) | 100 |
| LogicKor - Reasoning | 9.42 | Score (0-10) | 98.2 |
| JustEval - Clarity | 4.99 | Score (1-5) | 96.7 |
| JustEval - Factuality | 4.9 | Score (1-5) | 96.7 |
| LogicKor - Grammar | 8.35 | Score (0-10) | 96.5 |
| LiveBench Plot Unscrambling | 45.54 | Score | 95.8 |
| LiveBench Typos | 64 | Score | 95.8 |
| LogicKor - Math | 8.14 | Score (0-10) | 94.9 |
Interactive version: theaggregate.ai/model?slug=gpt-4-0613 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.