Lumimaid-v0.2-8B — benchmark results
NeverSleep's roleplay fine-tune of Llama 3.1 8B Instruct on a heavily cleaned RP dataset, the NSFW-capable successor to Lumimaid 0.1. Provider: Other. Released 2024-07-24. Access: Open.
Unified ELO 1467 ± 23, rank #924 of 1776 rated models, from 10 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - GPQA | 8.17 | Score | 68.6 |
| Open LLM Leaderboard - MuSR | 12.32 | Score | 67.2 |
| Open LLM Leaderboard - MATH Level 5 | 14.35 | Score | 60.8 |
| Open LLM Leaderboard - IFEval | 50.38 | Score | 59.6 |
| Open LLM Leaderboard - BBH | 31.96 | Score | 59.3 |
| Open LLM Leaderboard - MMLU-Pro | 29.29 | Score | 56.3 |
| UGI - Willingness (W/10) | 4.8 | W/10 Score | 36.1 |
| UGI - Writing | 28.33 | Writing Score | 31.9 |
| UGI Leaderboard | 24.5 | UGI Score | 18.5 |
| UGI - Natural Intelligence | 9.71 | NatInt Score | 6.5 |
Interactive version: theaggregate.ai/model?slug=lumimaid-v0-2-8b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.