DeepSeek V4 Flash 0731 (Reasoning, Max Effort)
DeepSeekDeepSeekओपन वेटMIT · व्यावसायिक उपयोग
विवरण
DeepSeek-V4-Flash-Max is the maximum reasoning effort mode of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window. Sharing the V4 series' hybrid attention architecture (Compressed Sparse Attention combined with Heavily Compressed Attention), Manifold-Constrained Hyper-Connections, and Muon optimizer, V4-Flash-Max delivers reasoning performance comparable to V4-Pro when given a larger thinking budget while operating at a fraction of the parameter scale. It is pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation.
रिलीज़ तिथि
2026-07-31
पैरामीटर
284.0B
संदर्भ लंबाई
1.0M
मोडैलिटीज़
text
क्षमता रडार
49
general
66
coding
91
reasoning
66
science
60
agents
0
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| एजेंटिक क्षमता | 74 | 45.0 | LS |
| कोडिंग रैंकिंग | 35 | 88.0 | AA |
| सामान्य रैंकिंग | 34 | 82.0 | AA |
| विज्ञान | 47 | 82.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Agents
Terminal-Bench 2.1
82.7%स्वयं
CyberGym
76.7%स्वयं
BrowseCompOpenAI (2025)
73.2%स्वयं
MCP Atlas
69.0%स्वयं
DSBench-FullStack
68.7%स्वयं
DSBench-Hard
59.6%स्वयं
Terminal-Bench 2.0Stanford × Laude Institute (2026)
56.9%स्वयं
DeepSWE
54.4%स्वयं
NL2Repo
54.2%स्वयं
SWE-Bench ProPrinceton NLP (2024)
52.6%स्वयं
Toolathlon
47.8%स्वयं
Agents' Last Exam
25.2%स्वयं
AutomationBench
25.1%स्वयं
Biology
GPQANYU + Cohere + Anthropic (2023)
88.1%स्वयं
Code
LiveCodeBench
91.6%स्वयं
SWE-Bench Verified
79.0%स्वयं
SWE-bench Multilingual
73.3%स्वयं
Factuality
SimpleQA
34.1%स्वयं
Finance
MMLU-Pro
86.2%स्वयं
General
CSimpleQA
78.9%स्वयं
MRCR 1M
78.7%स्वयं
CorpusQA 1M
60.5%स्वयं
Math
CodeForces
1.00 / 3000स्वयं
HMMT Feb 26
94.8%स्वयं
IMO-AnswerBench
88.4%स्वयं
MathArena Apex
85.7%स्वयं
Humanity's Last Exam
45.1%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Coding Index(Artificial Analysis)69.1
Intelligence Index(Artificial Analysis)51.8
Gpqa(NYU + Cohere + Anthropic (2023))0.9
Terminalbench V2 10.8
Lcr(Artificial Analysis)0.7
Scicode(UIUC + Argonne National Lab (2024))0.5
Tau Banking0.4
Hle(Center for AI Safety + Scale AI (2025))0.4
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Legal90
Physics90
Finance90
Healthcare90
Biology90
Chemistry90
Math80
Language80
Frontend Development80
Long Context70
Reasoning70
Search70
General70
Code70
Agents60
Tool Calling60
Vision50
Factuality30
मूल्य निर्धारण
इनपुट मूल्य$0.14 / 1M टोकन
आउटपुट मूल्य$0.28 / 1M टोकन
मिश्रित मूल्य (3:1)$0.175 / 1M टोकन
कैश पठन मूल्य$0.0028 / 1M टोकन
गति
टोकन/सेकंड109.1
पहले टोकन में देरी0.96s
पहले उत्तर में देरी19.29s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
8 प्रदाता
सबसे सस्ता: DeepSeekसबसे महंगा: TensorX
प्रदाताइनपुटआउटपुट
1DeepSeekसबसे सस्ता
$0
$0
2OpenRouter
$0.08
$0.18
3NanoGPT
$0.14
$0.28
4Kilo Gateway
$0.14
$0.28
5Ambient
$0.14
$0.28
6Merge Gateway
$0.14
$0.28
7Vercel AI Gateway
$0.2
$0.4
8TensorX
$0.25
$0.3
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।