DeepSeek R1 Distill Llama 8B — benchmark results
Provider: DeepSeek. Released 2025-01-20. Access: Open.
Unified ELO 1351 ± 15, rank #1407 of 1776 rated models, from 241 benchmark results.
Strongest benchmark results
| Benchmark | Score | Metric | Percentile |
|---|---|---|---|
| Open CoT - LogiQA | 13.26 | CoT Gain (%) | 99.2 |
| Open CoT - LogiQA 2 | 18.26 | CoT Gain (%) | 98.5 |
| Open CoT Leaderboard | 16.1 | Average CoT Gain (%) | 96.9 |
| Open CoT - LSAT Reading Comprehension | 22.68 | CoT Gain (%) | 93.9 |
| Open CoT - LSAT Analytical Reasoning | 8.26 | CoT Gain (%) | 89.3 |
| Open CoT - LSAT Logical Reasoning | 18.04 | CoT Gain (%) | 84.7 |
| FACTS Leaderboard | 40.68 | Combined Score (%) | 75.8 |
| Open LLM Leaderboard - MATH Level 5 | 21.98 | Score | 74.5 |
| HELM SeaHELM - Flores (id-en) | 56.47 | ChrF++ | 65 |
| Open FinLLM Reasoning - XBRL-Math | 81.11 | Accuracy (%) | 64 |
| LatamBoard - Spanish PAWS | 60.95 | Score (%) | 60.6 |
| Open Japanese LLM - Wiki NER SET F1 | 6.19 | Score (%) | 57.2 |
Interactive version: theaggregate.ai/model?slug=deepseek-r1-distill-llama-8b · How the rankings work · Data refreshed daily, snapshot 2026-07-22.