falcon-7B Instruct: benchmark results

Provider: TII. Released 2023-05-25. Access: Open.

Unified ELO 1258 ± 1, rank #1382 of 1392 rated models, from 98 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
LLM Trustworthy - Fairness100Trust Score (%)94
MMLU-by-task - TruthfulQA MC128.89Accuracy (%)52
MMLU-by-task - Machine Learning32.14Accuracy (%)48.7
MMLU-by-task - TruthfulQA MC244.07Accuracy (%)46.6
LLM Trustworthy - Stereotype87Trust Score (%)46
MMLU-by-task - Moral Scenarios25.14Accuracy (%)45.4
MMLU-by-task - Abstract Algebra29Accuracy (%)45.3
LLM Trustworthy - Ethics50.28Trust Score (%)44
MMLU-by-task - ARC Challenge42.15Accuracy (%)36.7
LLM Trustworthy - Adversarial43.98Trust Score (%)36
LLM Trustworthy - Toxicity39Trust Score (%)36
MMLU-by-task - Formal Logic26.98Accuracy (%)35.6

Interactive version: theaggregate.ai/model?slug=falcon-7b-instruct · How It Works · Data refreshed daily, snapshot 2026-09-05.