orca-mini-v5.8B-orpo: benchmark results

Provider: Other. Access: Open.

Unified ELO 1469 ± 20, rank #1567 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - GSM8K65.96Accuracy (%) (5-shot)81
Open LLM Leaderboard v1 - MMLU64.67Accuracy (%) (5-shot)74.6
Open LLM Leaderboard v1 - TruthfulQA MC253.44MC2 (%) (0-shot)59.5
Open LLM Leaderboard - MuSR41.31Score53
Open LLM Leaderboard - BBH49.64Score46.9
Open LLM Leaderboard v1 - HellaSwag79.93Normalized accuracy (%) (10-shot)39.6
Open LLM Leaderboard - GPQA28.44Score39.4
Open LLM Leaderboard v1 - WinoGrande74.82Accuracy (%) (5-shot)39
Open LLM Leaderboard v1 - ARC Challenge57.08Normalized accuracy (%) (25-shot)36.6
Open LLM Leaderboard - MATH Level 56.65Score35.9
Open LLM Leaderboard - MMLU-Pro29.47Score34.8
Open LLM Leaderboard - IFEval8.24Score1.4

Interactive version: theaggregate.ai/model?slug=orca-mini-v5-8b-orpo · How It Works · Data refreshed daily, snapshot 2026-09-23.