text-davinci-001 — benchmark results
InstructGPT completion model snapshot from the text-davinci series. Provider: OpenAI. Released 2022-01-27. Access: API.
Unified ELO 1279 ± 47, rank #1576 of 1776 rated models, from 15 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Epoch AI - Lambada | 86.4 | Score | 96 |
| Epoch AI - Science Qa | 74.04 | Score | 73.9 |
| OpenBookQA | 65.4 | Accuracy (%) | 73.8 |
| WinoGrande | 77.7 | Accuracy (%) | 66.2 |
| PIQA | 82.3 | Accuracy (%) | 65 |
| HellaSwag | 79.3 | Accuracy (%) | 61.8 |
| ARC Challenge (AI2) | 53.2 | Accuracy (%) | 44.2 |
| BoolQ | 77.5 | Accuracy (%) | 41.6 |
| TriviaQA | 71.2 | Accuracy (%) | 25.6 |
| AlpacaEval 2.0 | 9.03 | LC Win Rate (%) | 18.9 |
| Epoch AI - Common Sense Qa 2 | 52.9 | Score | 16.7 |
| MMLU | 43.9 | Accuracy (%) | 15.6 |
Interactive version: theaggregate.ai/model?slug=text-davinci-001 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.