DeepSeek V4 Flash (Reasoning) — benchmark results

Provider: DeepSeek. Released 2026-04-23. Access: Open.

Unified ELO 1712 ± 32, rank #254 of 1776 rated models, from 8 benchmark results.

Strongest benchmark results

BenchmarkScoreMetricPercentile
UGI Leaderboard56.63UGI Score97.4
Wolfram LLM Benchmarking Project59.6Correct Functionality (%)89.3
UGI - Natural Intelligence45.52NatInt Score89
UGI - Writing49.74Writing Score88.9
GRIPS88.2Accuracy (%)61.5
GRIPS-hard49.3Accuracy (%)61.5
UGI - Willingness (W/10)5.2W/10 Score41.4
SpeechMap Compliance50.7% Requests Completed32.1

Interactive version: theaggregate.ai/model?slug=deepseek-v4-flash-reasoning · How the rankings work · Data refreshed daily, snapshot 2026-07-22.