DeepSeek R1 0528 Qwen3 8B: benchmark results

DeepSeek's R1-0528 reasoning distilled into a Qwen3 8B dense student model. Provider: DeepSeek. Released 2025-05-28. Access: Open.

Unified ELO 1481 ± 1, rank #779 of 1392 rated models, from 151 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
HardcoreLogic - Unsolvable Puzzles95.19Accuracy (%)85
EuroEval Danish NLU - Dansk64.18Named entity recognition Score (%)83.5
EuroEval English NLU - CoNLL EN79.02Named entity recognition Score (%)80.9
FACTS Leaderboard41.1Combined Score (%)78.8
EuroEval French NLU - Eltec65.69Named entity recognition Score (%)77.8
AA Omniscience - Software Engineering (SWE) - HTML40Accuracy (%)75.9
EuroEval Portuguese NLU - HAREM49.67Named entity recognition Score (%)75.8
EuroEval Norwegian NLU - NorNE NN72.03Named entity recognition Score (%)75.1
EuroEval Italian NLU - MultiNERD IT72.4Named entity recognition Score (%)74.7
EuroEval Norwegian NLU - NorNE NB73.66Named entity recognition Score (%)74.4
EuroEval Italian Knowledge63.79Knowledge Average Score (%)74.1
EuroEval Spanish NLU - CoNLL ES69.5Named entity recognition Score (%)73.9

Interactive version: theaggregate.ai/model?slug=deepseek-r1-0528-qwen3-8b · How It Works · Data refreshed daily, snapshot 2026-09-05.