CatunaMayo: benchmark results

Provider: Other. Access: Open.

Unified ELO 1505 ± 20, rank #1159 of 2928 rated models, from 12 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open LLM Leaderboard v1 - GSM8K72.33Accuracy (%) (5-shot)97.7
Open LLM Leaderboard v1 - ARC Challenge71.76Normalized accuracy (%) (25-shot)90.2
Open LLM Leaderboard v1 - TruthfulQA MC269.96MC2 (%) (0-shot)89.7
Open LLM Leaderboard v1 - HellaSwag87.9Normalized accuracy (%) (10-shot)88.6
Open LLM Leaderboard - MuSR45.4Score85.3
Open LLM Leaderboard v1 - WinoGrande82.56Accuracy (%) (5-shot)84.6
Open LLM Leaderboard v1 - MMLU65.21Accuracy (%) (5-shot)83
Open LLM Leaderboard - BBH52.44Score61.4
Open LLM Leaderboard - GPQA29.19Score46.9
Open LLM Leaderboard - MMLU-Pro31.78Score43.1
Open LLM Leaderboard - MATH Level 58.46Score42.5
Open LLM Leaderboard - IFEval40.74Score40.9

Interactive version: theaggregate.ai/model?slug=catunamayo · How It Works · Data refreshed daily, snapshot 2026-09-23.