OLMo-7B-hf: benchmark results

Provider: Allen AI. Access: Open.

Unified ELO 1272 ± 22, rank #2773 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - HellaSwag77.31Normalized accuracy (%) (10-shot)30.5
Open LLM Leaderboard - GPQA27.27Score27.6
Open LLM Leaderboard - IFEval27.19Score24.2
Open LLM Leaderboard v1 - WinoGrande69.38Accuracy (%) (5-shot)23.4
Open LLM Leaderboard v1 - GSM8K3.79Accuracy (%) (5-shot)22
Open LLM Leaderboard v1 - ARC Challenge45.65Normalized accuracy (%) (25-shot)21
Open LLM Leaderboard v1 - MMLU28.13Accuracy (%) (5-shot)15.7
Open LLM Leaderboard - BBH32.79Score15.2
Open LLM Leaderboard - MuSR34.87Score12
Open LLM Leaderboard - MATH Level 51.21Score8.7
Open LLM Leaderboard - MMLU-Pro11.73Score8.6
Open LLM Leaderboard v1 - TruthfulQA MC235.93MC2 (%) (0-shot)2.9

Interactive version: theaggregate.ai/model?slug=olmo-7b-hf · How It Works · Data refreshed daily, snapshot 2026-09-23.