FuseChat-Llama-3.1-8B-Instruct — benchmark results
FuseAI's implicit model-fusion tune of Llama 3.1 8B Instruct, distilling Gemma 2 27B, Mistral Large, Qwen2.5 72B and Llama 3.1 70B via SFT+DPO. Provider: Other. Released 2024-11-20. Access: Open.
Unified ELO 1537 ± 62, rank #657 of 1776 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| AlpacaEval 2.0 | 65.39 | LC Win Rate (%) | 94.6 |
| Open LLM Leaderboard - IFEval | 72.05 | Score | 85.7 |
| Open LLM Leaderboard - MATH Level 5 | 24.77 | Score | 77.9 |
| Open LLM Leaderboard - GPQA | 7.38 | Score | 62.2 |
| Open LLM Leaderboard - MMLU-Pro | 30.37 | Score | 61.7 |
| Open LLM Leaderboard - BBH | 30.85 | Score | 54.7 |
| Open LLM Leaderboard - MuSR | 6.15 | Score | 29.5 |
Interactive version: theaggregate.ai/model?slug=fusechat-llama-3-1-8b-instruct · How the rankings work · Data refreshed daily, snapshot 2026-07-22.