OPT-175B: benchmark results
Provider: Meta. Released 2022-05-03. Access: Open.
Unified ELO 1310 ± 34, rank #2421 of 2656 rated models, from 36 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| HELM Classic - XSUM | 15.49 | ROUGE-2 (%) | 92.7 |
| HELM Classic - OpenbookQA | 58.6 | Exact Match (%) | 88.7 |
| HELM Classic - HellaSwag | 79.1 | Exact Match (%) | 77.4 |
| HELM Classic - LSAT | 22.03 | Exact Match (%) | 77.2 |
| HELM Classic - Synthetic Reasoning Natural | 24.79 | F1 (%) | 76.5 |
| HELM Classic - MS MARCO Regular | 28.78 | RR@10 (%) | 72.4 |
| HELM Classic - BLiMP | 83.05 | Exact Match (%) | 71 |
| HELM Classic - BoolQ | 79.3 | Exact Match (%) | 69.7 |
| HELM Classic - IMDB | 94.73 | Exact Match (%) | 68.9 |
| HELM Classic - CNN/DailyMail | 14.59 | ROUGE-2 (%) | 65.9 |
| HELM Classic - bAbI | 50.66 | Exact Match (%) | 65.2 |
| HELM | 60.95 | Mean win rate (self-reported) | 61.5 |
Interactive version: theaggregate.ai/model?slug=opt-175b · How It Works · Data refreshed daily, snapshot 2026-09-19.