メインコンテンツへスキップ

DeepSeek V4 Flash 0731 (Reasoning, Max Effort)

DeepSeekDeepSeekオープンウエイトMIT · 商用利用可

説明

DeepSeek-V4-Flash-Max is the maximum reasoning effort mode of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window. Sharing the V4 series' hybrid attention architecture (Compressed Sparse Attention combined with Heavily Compressed Attention), Manifold-Constrained Hyper-Connections, and Muon optimizer, V4-Flash-Max delivers reasoning performance comparable to V4-Pro when given a larger thinking budget while operating at a fraction of the parameter scale. It is pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation.

リリース日
2026-07-31
パラメータ
284.0B
コンテキスト長
1.0M
モダリティ
text

能力レーダー

49
general
66
coding
91
reasoning
66
science
60
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
エージェント能力74
45.0
LS
コーディングランキング35
88.0
AA
総合ランキング34
82.0
AA
科学47
82.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

Terminal-Bench 2.182.7%自己申告
CyberGym76.7%自己申告
BrowseCompOpenAI (2025)73.2%自己申告
MCP Atlas69.0%自己申告
DSBench-FullStack68.7%自己申告
DSBench-Hard59.6%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)56.9%自己申告
DeepSWE54.4%自己申告
NL2Repo54.2%自己申告
SWE-Bench ProPrinceton NLP (2024)52.6%自己申告
Toolathlon47.8%自己申告
Agents' Last Exam25.2%自己申告
AutomationBench25.1%自己申告

Biology

GPQANYU + Cohere + Anthropic (2023)88.1%自己申告

Code

LiveCodeBench91.6%自己申告
SWE-Bench Verified79.0%自己申告
SWE-bench Multilingual73.3%自己申告

Factuality

SimpleQA34.1%自己申告

Finance

MMLU-Pro86.2%自己申告

General

CSimpleQA78.9%自己申告
MRCR 1M78.7%自己申告
CorpusQA 1M60.5%自己申告

Math

CodeForces1.00 / 3000自己申告
HMMT Feb 2694.8%自己申告
IMO-AnswerBench88.4%自己申告
MathArena Apex85.7%自己申告
Humanity's Last Exam45.1%自己申告

AA評価指数

(Artificial Analysis)
Coding Index(Artificial Analysis)
69.1
Intelligence Index(Artificial Analysis)
51.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Terminalbench V2 1
0.8
Lcr(Artificial Analysis)
0.7
Scicode(UIUC + Argonne National Lab (2024))
0.5
Tau Banking
0.4
Hle(Center for AI Safety + Scale AI (2025))
0.4

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Legal
90
Physics
90
Finance
90
Healthcare
90
Biology
90
Chemistry
90
Math
80
Language
80
Frontend Development
80
Long Context
70
Reasoning
70
Search
70
General
70
Code
70
Agents
60
Tool Calling
60
Vision
50
Factuality
30

価格設定

入力価格$0.14 / 1Mトークン
出力価格$0.28 / 1Mトークン
混合価格(3:1)$0.175 / 1Mトークン
キャッシュ読み取り価格$0.0028 / 1Mトークン

速度

トークン/秒109.1
初トークン遅延0.96s
初回答遅延19.29s

プロバイダー価格ランキング

プロバイダー価格ランキング

8 プロバイダー

最安: DeepSeek最高: TensorX
プロバイダー入力出力
1DeepSeek最安
$0
$0
2OpenRouter
$0.08
$0.18
3NanoGPT
$0.14
$0.28
4Kilo Gateway
$0.14
$0.28
5Ambient
$0.14
$0.28
6Merge Gateway
$0.14
$0.28
7Vercel AI Gateway
$0.2
$0.4
8TensorX
$0.25
$0.3

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク