Evil-Alpaca-3B-L3.2 — benchmark results
Provider: Other. Released 2024-09-28. Access: Open.
Unified ELO 1362 ± 37, rank #1374 of 1776 rated models, from 10 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 10.94 | Score | 56 |
| Open LLM Leaderboard - MATH Level 5 | 7.02 | Score | 37.5 |
| UGI - Willingness (W/10) | 4.5 | W/10 Score | 33.7 |
| Open LLM Leaderboard - BBH | 20.85 | Score | 31.2 |
| Open LLM Leaderboard - IFEval | 32.51 | Score | 30.4 |
| Open LLM Leaderboard - MMLU-Pro | 18.01 | Score | 27.5 |
| UGI Leaderboard | 24.57 | UGI Score | 18.7 |
| Open LLM Leaderboard - GPQA | 1.79 | Score | 17.3 |
| UGI - Natural Intelligence | 9.86 | NatInt Score | 6.8 |
| UGI - Writing | 11.9 | Writing Score | 3.3 |
Interactive version: theaggregate.ai/model?slug=evil-alpaca-3b-l3-2 · How the rankings work · Data refreshed daily, snapshot 2026-07-22.