Llama 2 13B Chat Gptq: benchmark results

Provider: Meta. Released 2023-07-18. Access: Open.

Unified ELO 1428 ± 46, rank #1685 of 2656 rated models, from 10 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLM Trustworthy - Privacy98.88Trust Score (%)100
LLM Trustworthy - Stereotype100Trust Score (%)93.8
LLM Trustworthy - Toxicity80.87Trust Score (%)87.5
AI-Secure LLM Trustworthy Leaderboard0.72Trustworthy average70.8
LLM Trustworthy Leaderboard71.99Average Trust Score (%)70.8
LLM Trustworthy - Adversarial Demo67.2Trust Score (%)62.5
LLM Trustworthy - Ethics53.93Trust Score (%)45.8
LLM Trustworthy - Fairness89.67Trust Score (%)45.8
LLM Trustworthy - Out-of-Distribution59.1Trust Score (%)31.2
LLM Trustworthy - Adversarial37.12Trust Score (%)4.2

Interactive version: theaggregate.ai/model?slug=llama-2-13b-chat-gptq · How It Works · Data refreshed daily, snapshot 2026-09-19.