DeepHermes 3 - Mistral 24B Preview (Non-reasoning) — benchmark results
Provider: Nous Research. Released 2025-01-15. Access: Open.
Unified ELO 1403 ± 48, rank #1257 of 1841 rated models, from 7 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Epoch AI - Scicode | 22.8 | Score | 23.3 |
| AA MMLU-Pro | 58.01 | Accuracy (%) | 21.5 |
| AA LiveCodeBench | 19.47 | Pass@1 (%) | 20.6 |
| AA MATH-500 | 59.47 | Accuracy (%) | 20.4 |
| Artificial Analysis Intelligence Index | 5.31 | Intelligence Index | 16.4 |
| AA GPQA Diamond | 38.18 | Accuracy (%) | 15.2 |
| AA Humanity's Last Exam | 3.92 | Accuracy (%) | 10.7 |
Interactive version: theaggregate.ai/model?slug=deephermes-3-mistral-24b-preview-non-reasoning · How It Works · Data refreshed daily, snapshot 2026-07-25.