Llama 2 Chat 7B — benchmark results
Provider: Meta. Released 2023-07-18. Access: Open.
Unified ELO 922 ± 151, rank #1839 of 1841 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA Humanity's Last Exam | 5.81 | Accuracy (%) | 42.3 |
| Artificial Analysis Intelligence Index | 4.26 | Intelligence Index | 12.2 |
| AA GPQA Diamond | 22.73 | Accuracy (%) | 1.3 |
| AA MMLU-Pro | 16.39 | Accuracy (%) | 1.2 |
| Epoch AI - Scicode | 0 | Score | 0.2 |
| AA LiveCodeBench | 0.21 | Pass@1 (%) | 0 |
| AA MATH-500 | 5.87 | Accuracy (%) | 0 |
Interactive version: theaggregate.ai/model?slug=llama-2-chat-7b · How It Works · Data refreshed daily, snapshot 2026-07-25.