DeepSeek V4 Flash 0731 (Reasoning, Max Effort)
DeepSeekDeepSeekOpen WeightMIT · Commercial OK
Description
DeepSeek-V4-Flash-Max is the maximum reasoning effort mode of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window. Sharing the V4 series' hybrid attention architecture (Compressed Sparse Attention combined with Heavily Compressed Attention), Manifold-Constrained Hyper-Connections, and Muon optimizer, V4-Flash-Max delivers reasoning performance comparable to V4-Pro when given a larger thinking budget while operating at a fraction of the parameter scale. It is pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation.
Release Date
2026-07-31
Parameters
284.0B
Context Length
1.0M
Modalities
text
Capability Radar
49
general
66
coding
91
reasoning
66
science
60
agents
0
multimodal
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 74 | 45.0 | LS |
| Code Ranking | 35 | 88.0 | AA |
| General Ranking | 34 | 82.0 | AA |
| Science | 47 | 82.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
Terminal-Bench 2.1
82.7%SR
CyberGym
76.7%SR
BrowseCompOpenAI (2025)
73.2%SR
MCP Atlas
69.0%SR
DSBench-FullStack
68.7%SR
DSBench-Hard
59.6%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
56.9%SR
DeepSWE
54.4%SR
NL2Repo
54.2%SR
SWE-Bench ProPrinceton NLP (2024)
52.6%SR
Toolathlon
47.8%SR
Agents' Last Exam
25.2%SR
AutomationBench
25.1%SR
Biology
GPQANYU + Cohere + Anthropic (2023)
88.1%SR
Code
LiveCodeBench
91.6%SR
SWE-Bench Verified
79.0%SR
SWE-bench Multilingual
73.3%SR
Factuality
SimpleQA
34.1%SR
Finance
MMLU-Pro
86.2%SR
General
CSimpleQA
78.9%SR
MRCR 1M
78.7%SR
CorpusQA 1M
60.5%SR
Math
CodeForces
1.00 / 3000SR
HMMT Feb 26
94.8%SR
IMO-AnswerBench
88.4%SR
MathArena Apex
85.7%SR
Humanity's Last Exam
45.1%SR
AA Evaluation Indices
(Artificial Analysis)Coding Index(Artificial Analysis)69.1
Intelligence Index(Artificial Analysis)51.8
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.8
Lcr(Artificial Analysis)0.7
Scicode(UIUC + Argonne National Lab (2024))0.5
Tau Banking0.4
Hle(Center for AI Safety + Scale AI (2025))0.4
LLM Stats Category Scores
(LLM Stats (zeroeval))Legal90
Physics90
Finance90
Healthcare90
Biology90
Chemistry90
Math80
Language80
Frontend Development80
Long Context70
Reasoning70
Search70
General70
Code70
Agents60
Tool Calling60
Vision50
Factuality30
Pricing
Input Price$0.14 / 1M tokens
Output Price$0.28 / 1M tokens
Blended Price (3:1)$0.175 / 1M tokens
Cache Read Price$0.0028 / 1M tokens
Speed
Tokens/sec109.1
Time to First Token0.96s
Time to Answer19.29s
Provider Price Ranking
Provider Price Ranking
8 providers
Cheapest: DeepSeekMost Expensive: TensorX
ProviderInputOutput
1DeepSeekCheapest
$0
$0
2OpenRouter
$0.08
$0.18
3NanoGPT
$0.14
$0.28
4Kilo Gateway
$0.14
$0.28
5Ambient
$0.14
$0.28
6Merge Gateway
$0.14
$0.28
7Vercel AI Gateway
$0.2
$0.4
8TensorX
$0.25
$0.3
Compare pricing across different API providers for this model.