Skip to main content

DeepSeek-V4-Flash-0423

DeepSeekDeepSeekOpen WeightMIT · Commercial OK

Description

DeepSeek-V4-Flash-0423 is the preview release of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window, evaluated here at the default high reasoning effort. It shares the V4 series' hybrid attention architecture combining Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA) for dramatically improved long-context efficiency, Manifold-Constrained Hyper-Connections (mHC) for stable signal propagation, and the Muon optimizer for faster convergence. Pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation, V4-Flash offers reasoning capabilities that closely approach V4-Pro with faster responses and highly cost-effective pricing.

Release Date
2026-04-23
Parameters
284.0B
Context Length
Modalities
text

Capability Radar

70
general
70
coding
80
reasoning
77
scienceest.
60
agents
0
multimodal

Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.

Rankings

Domain#RankScoreSource
Agentic Capability117
35.0
LS

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

MCP Atlas67.4%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)56.6%SR
BrowseCompOpenAI (2025)53.5%SR
SWE-Bench ProPrinceton NLP (2024)52.3%SR
Toolathlon43.5%SR

Biology

GPQANYU + Cohere + Anthropic (2023)87.4%SR

Code

LiveCodeBench88.4%SR
SWE-Bench Verified78.6%SR
SWE-bench Multilingual70.2%SR

Factuality

SimpleQA28.9%SR

Finance

MMLU-Pro86.4%SR

General

MRCR 1M76.9%SR
CSimpleQA73.2%SR
CorpusQA 1M59.3%SR

Math

CodeForces0.94 / 3000SR
HMMT Feb 2691.9%SR
IMO-AnswerBench85.1%SR
MathArena Apex72.1%SR
Humanity's Last Exam40.3%SR

AA Evaluation Indices

(Artificial Analysis)

No AA evaluation data available

LLM Stats Category Scores

(LLM Stats (zeroeval))
Legal
90
Physics
90
Finance
90
Healthcare
90
Biology
90
Chemistry
90
Math
80
Language
80
Frontend Development
80
Long Context
70
Reasoning
70
General
70
Code
70
Tool Calling
60
Search
50
Agents
50
Vision
40
Factuality
30

Pricing

Input Price$0 / 1M tokens
Output Price$0 / 1M tokens
Blended Price (3:1)$0 / 1M tokens

Speed

No speed data available

Provider Price Ranking

Provider Price Ranking

3 providers

Cheapest: DeepSeekMost Expensive: Novita
ProviderInputOutput
1DeepSeekPRIMARY
$0
$0
2DeepInfra
$0
$0
3Novita
$0
$0

Compare pricing across different API providers for this model.

External Sources