ROGERphi-7B-slerp: benchmark results

Provider: Other. Access: Open.

Unified ELO 1490 ± 20, rank #1355 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard - MuSR46.85Score90.7
Open LLM Leaderboard v1 - GSM8K68.69Accuracy (%) (5-shot)86.8
Open LLM Leaderboard v1 - HellaSwag86.19Normalized accuracy (%) (10-shot)79.4
Open LLM Leaderboard v1 - ARC Challenge67.58Normalized accuracy (%) (25-shot)77.4
Open LLM Leaderboard v1 - TruthfulQA MC259.84MC2 (%) (0-shot)74.6
Open LLM Leaderboard v1 - WinoGrande80.11Accuracy (%) (5-shot)72.4
Open LLM Leaderboard v1 - MMLU64.15Accuracy (%) (5-shot)66.2
Open LLM Leaderboard - BBH51.96Score58.9
Open LLM Leaderboard - GPQA28.86Score43.2
Open LLM Leaderboard - MATH Level 57.33Score38.6
Open LLM Leaderboard - MMLU-Pro30.53Score38.2
Open LLM Leaderboard - IFEval38.61Score38

Interactive version: theaggregate.ai/model?slug=rogerphi-7b-slerp · How It Works · Data refreshed daily, snapshot 2026-09-23.