Mahou-1.2a-mistral-7B: benchmark results

Provider: Other. Access: Open.

Unified ELO 1485 ± 19, rank #1409 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - HellaSwag88.14Normalized accuracy (%) (10-shot)89.7
Open LLM Leaderboard v1 - TruthfulQA MC268.71MC2 (%) (0-shot)88
Open LLM Leaderboard v1 - ARC Challenge68.77Normalized accuracy (%) (25-shot)81.4
Open LLM Leaderboard v1 - GSM8K63.61Accuracy (%) (5-shot)77.2
Open LLM Leaderboard v1 - MMLU64.73Accuracy (%) (5-shot)75.6
Open LLM Leaderboard v1 - WinoGrande80.11Accuracy (%) (5-shot)72.4
Open LLM Leaderboard - BBH51.18Score54.8
Open LLM Leaderboard - IFEval45.52Score50.5
Open LLM Leaderboard - MMLU-Pro31.63Score42.6
Open LLM Leaderboard - MATH Level 56.87Score36.9
Open LLM Leaderboard - MuSR38.96Score34.9
Open LLM Leaderboard - GPQA27.18Score26.7

Interactive version: theaggregate.ai/model?slug=mahou-1-2a-mistral-7b · How It Works · Data refreshed daily, snapshot 2026-09-23.