Llama 2 13B Base — benchmark results
Provider: Meta. Released 2023-07-18. Access: Open.
Unified ELO 1320 ± 22, rank #1492 of 1776 rated models, from 113 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open Japanese LLM - Xlsum JA Bleu JA | 75.69 | Score (%) | 99.3 |
| EuroEval Portuguese NLU - MultiWikiQA PT | 76.05 | Reading comprehension Score (%) | 88.9 |
| EuroEval Italian NLU - SQuAD IT | 71.05 | Reading comprehension Score (%) | 76.8 |
| EuroEval Dutch NLU - DBRD | 90.19 | Sentiment classification Score (%) | 72.6 |
| EuroEval Spanish NLU - MLQA ES | 63.41 | Reading comprehension Score (%) | 72.4 |
| EuroEval Portuguese NLU - SST-2 PT | 80.59 | Sentiment classification Score (%) | 70.5 |
| Open Japanese LLM - Wiki Dependency SET F1 | 36.98 | Score (%) | 68 |
| URIAL-Bench - Writing | 7.53 | Judge Score (0-10) | 61.1 |
| OpenEval - BBQ | 61.66 | Exact Match (%) | 59.5 |
| EuroEval Portuguese NLU | 50.31 | NLU Average Score (%) | 52.7 |
| URIAL-Bench - Extraction | 4.7 | Judge Score (0-10) | 50 |
| EuroEval Danish Knowledge | 46.76 | Knowledge Average Score (%) | 47.7 |
Interactive version: theaggregate.ai/model?slug=llama-2-13b-base · How the rankings work · Data refreshed daily, snapshot 2026-07-22.