NeuralBeagle14-7B — benchmark results
mlabonne's DPO fine-tune of his Beagle14-7B Mistral merge, briefly the top-ranked 7B on the Open LLM Leaderboard in January 2024. Provider: Other. Released 2024-01-15. Access: Open.
Unified ELO 1448 ± 23, rank #1006 of 1776 rated models, from 36 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| EuroEval German NLU - Sb10K | 59.6 | Sentiment classification Score (%) | 93.1 |
| EuroEval Dutch NLU - CoNLL NL | 63.53 | Named entity recognition Score (%) | 73.7 |
| EuroEval German NLU - GermEval | 64.81 | Named entity recognition Score (%) | 73 |
| Open LLM Leaderboard - MuSR | 12.89 | Score | 70.7 |
| EuroEval English Common Sense Reasoning | 71.96 | Common Sense Reasoning Average Score (%) | 69.7 |
| EuroEval Swedish NLU - SUC3 | 61.25 | Named entity recognition Score (%) | 62.7 |
| EuroEval Swedish NLU - Swerec | 76.03 | Sentiment classification Score (%) | 60.4 |
| EuroEval Dutch Common Sense Reasoning | 47.87 | Common Sense Reasoning Average Score (%) | 59.4 |
| Open LLM Leaderboard - IFEval | 49.35 | Score | 57.8 |
| EuroEval Icelandic NLU - MIM-GOLD NER | 49.86 | Named entity recognition Score (%) | 56.8 |
| EuroEval Swedish Common Sense Reasoning | 38.78 | Common Sense Reasoning Average Score (%) | 54.9 |
| EuroEval German Common Sense Reasoning | 49.13 | Common Sense Reasoning Average Score (%) | 54.7 |
Interactive version: theaggregate.ai/model?slug=neuralbeagle14-7b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.