DeepSeek-V4-Flash-0423
説明
DeepSeek-V4-Flash-0423 is the preview release of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window, evaluated here at the default high reasoning effort. It shares the V4 series' hybrid attention architecture combining Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA) for dramatically improved long-context efficiency, Manifold-Constrained Hyper-Connections (mHC) for stable signal propagation, and the Muon optimizer for faster convergence. Pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation, V4-Flash offers reasoning capabilities that closely approach V4-Pro with faster responses and highly cost-effective pricing.
能力レーダー
専用の科学ベンチマークがない場合、Science は LLM Stats の科学スコアまたは推論能力から推定します。
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 117 | 35.0 | LS |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
Biology
Code
Factuality
Finance
General
Math
AA評価指数
(Artificial Analysis)AA評価データがありません
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))価格設定
速度
速度データがありません
プロバイダー価格ランキング
プロバイダー価格ランキング
3 プロバイダー
このモデルの異なるAPIプロバイダー間の価格を比較。