RolePlayLake-7B: benchmark results

Provider: Other. Access: Open.

Unified ELO 1498 ± 20, rank #1253 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - WinoGrande83.27Accuracy (%) (5-shot)87.9
Open LLM Leaderboard v1 - ARC Challenge70.56Normalized accuracy (%) (25-shot)86.4
Open LLM Leaderboard v1 - HellaSwag87.42Normalized accuracy (%) (10-shot)86.1
Open LLM Leaderboard v1 - TruthfulQA MC264.38MC2 (%) (0-shot)82.2
Open LLM Leaderboard - MuSR44.59Score81.3
Open LLM Leaderboard v1 - GSM8K65.05Accuracy (%) (5-shot)79.4
Open LLM Leaderboard v1 - MMLU64.55Accuracy (%) (5-shot)72.4
Open LLM Leaderboard - BBH52.52Score61.8
Open LLM Leaderboard - GPQA30.37Score60.3
Open LLM Leaderboard - IFEval50.57Score60
Open LLM Leaderboard - MMLU-Pro31.6Score42.5
Open LLM Leaderboard - MATH Level 57.25Score38.4

Interactive version: theaggregate.ai/model?slug=roleplaylake-7b · How It Works · Data refreshed daily, snapshot 2026-09-23.