magnum-v2-12B — benchmark results
Anthracite's magnum v2 fine-tune of Mistral Nemo Base 2407 (12B), aimed at replicating the prose quality of Claude 3 Sonnet and Opus. Provider: Other. Released 2024-08-03. Access: Open.
Unified ELO 1440 ± 29, rank #1043 of 1776 rated models, from 10 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open LLM Leaderboard - MuSR | 11.37 | Score | 59.4 |
| UGI - Writing | 32.97 | Writing Score | 48.3 |
| Open LLM Leaderboard - BBH | 28.79 | Score | 47.6 |
| Open LLM Leaderboard - GPQA | 5.48 | Score | 46 |
| Open LLM Leaderboard - MMLU-Pro | 24.08 | Score | 42.8 |
| Open LLM Leaderboard - IFEval | 37.62 | Score | 36.7 |
| Open LLM Leaderboard - MATH Level 5 | 5.44 | Score | 30.2 |
| UGI - Willingness (W/10) | 3.2 | W/10 Score | 24.1 |
| UGI - Natural Intelligence | 15.74 | NatInt Score | 22.5 |
| UGI Leaderboard | 21.54 | UGI Score | 14.2 |
Interactive version: theaggregate.ai/model?slug=magnum-v2-12b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.