dolphin-2.9.3-mistral-nemo-12B — benchmark results

Eric Hartford's uncensored Dolphin fine-tune of Mistral Nemo 12B with conversational, coding, and function-calling skills. Provider: Cognitive Computations. Released 2024-07-23. Access: Open.

Unified ELO 1503 ± 30, rank #788 of 1776 rated models, from 7 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
Open Korean LLM Leaderboard661.05Average Score (%)95
Open LLM Leaderboard - MuSR15.21Score82
Open LLM Leaderboard - BBH36.08Score73.7
Open LLM Leaderboard - GPQA8.72Score71.9
Open LLM Leaderboard - IFEval56.01Score67.9
Open LLM Leaderboard - MMLU-Pro26.41Score48.2
Open LLM Leaderboard - MATH Level 57.4Score38.9

Interactive version: theaggregate.ai/model?slug=dolphin-2-9-3-mistral-nemo-12b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.