Fireball-Alpaca-Llama3.1.07-8B-Philos-Math-KTO-beta — benchmark results
EpistemeAI's KTO-aligned Llama 3.1 8B Instruct fine-tune from its Fireball Alpaca line, targeting philosophy and math reasoning. Provider: Other. Released 2024-09-12. Access: Open.
Unified ELO 1428 ± 52, rank #1093 of 1776 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - IFEval | 72.74 | Score | 86.6 |
| Open Korean LLM Leaderboard | 490.55 | Average Score (%) | 86.5 |
| Open LLM Leaderboard - MATH Level 5 | 15.26 | Score | 62.4 |
| Open LLM Leaderboard - MMLU-Pro | 28.26 | Score | 52.9 |
| Open LLM Leaderboard - BBH | 26.9 | Score | 42.4 |
| Open LLM Leaderboard - GPQA | 4.03 | Score | 35.2 |
| Open LLM Leaderboard - MuSR | 4.28 | Score | 21 |
Interactive version: theaggregate.ai/model?slug=fireball-alpaca-llama3-1-07-8b-philos-math-kto-beta · How the rankings work · Data refreshed daily, snapshot 2026-07-22.