Llama 2 13B Chat AWQ: benchmark results

Provider: Meta. Released 2023-07-18. Access: Open.

Unified ELO 1431 ± 46, rank #1656 of 2656 rated models, from 10 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLM Trustworthy - Stereotype100Trust Score (%)93.8
LLM Trustworthy - Toxicity80.96Trust Score (%)91.7
LLM Trustworthy - Privacy96.31Trust Score (%)87.5
AI-Secure LLM Trustworthy Leaderboard0.71Trustworthy average66.7
LLM Trustworthy Leaderboard71.32Average Trust Score (%)66.7
LLM Trustworthy - Adversarial Demo66.29Trust Score (%)58.3
LLM Trustworthy - Ethics62.81Trust Score (%)54.2
LLM Trustworthy - Fairness78.19Trust Score (%)25
LLM Trustworthy - Out-of-Distribution58.38Trust Score (%)25
LLM Trustworthy - Adversarial41.99Trust Score (%)12.5

Interactive version: theaggregate.ai/model?slug=llama-2-13b-chat-awq · How It Works · Data refreshed daily, snapshot 2026-09-19.