Falcon3-7B-Base — benchmark results
Provider: TII. Released 2024-11-21. Access: Open.
Unified ELO 1375 ± 11, rank #1325 of 1776 rated models, from 204 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 18.14 | Score | 90.2 |
| Open LLM Leaderboard - GPQA | 12.86 | Score | 86.7 |
| EuroEval Spanish NLU - MLQA ES | 65.37 | Reading comprehension Score (%) | 84.2 |
| EuroEval Portuguese NLU - MultiWikiQA PT | 73.64 | Reading comprehension Score (%) | 76.7 |
| EuroEval Portuguese NLU - SST-2 PT | 81.44 | Sentiment classification Score (%) | 74.2 |
| Open LLM Leaderboard - MMLU-Pro | 32.34 | Score | 70.9 |
| Open LLM Leaderboard - MATH Level 5 | 19.41 | Score | 70 |
| EuroEval Spanish NLU | 45.92 | NLU Average Score (%) | 61.2 |
| EuroEval Portuguese | 48.27 | Average Score (%) | 57.6 |
| EuroEval Spanish NLU - Sentiment Headlines ES | 41.57 | Sentiment classification Score (%) | 57.5 |
| Open LLM Leaderboard - BBH | 31.56 | Score | 57.5 |
| EuroEval Portuguese NLU | 51.46 | NLU Average Score (%) | 57.3 |
Interactive version: theaggregate.ai/model?slug=falcon3-7b-base · How the rankings work · Data refreshed daily, snapshot 2026-07-22.