Llama 2 Chat 13B — benchmark results
Provider: Meta. Released 2023-07-18. Access: Open.
Unified ELO 1282 ± 72, rank #1630 of 1841 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AA Humanity's Last Exam | 4.73 | Accuracy (%) | 26.9 |
| Epoch AI - Scicode | 11.81 | Score | 11.3 |
| AA GPQA Diamond | 32.12 | Accuracy (%) | 9.6 |
| AA MMLU-Pro | 40.63 | Accuracy (%) | 8.1 |
| AA MATH-500 | 32.87 | Accuracy (%) | 7.8 |
| AA LiveCodeBench | 9.84 | Pass@1 (%) | 7.7 |
| Artificial Analysis Intelligence Index | 3 | Intelligence Index | 7.5 |
Interactive version: theaggregate.ai/model?slug=llama-2-chat-13b · How It Works · Data refreshed daily, snapshot 2026-07-25.