DeepSeek V4 Flash 0731 (Reasoning, Max Effort)
Description
DeepSeek-V4-Flash-Max is the maximum reasoning effort mode of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window. Sharing the V4 series' hybrid attention architecture (Compressed Sparse Attention combined with Heavily Compressed Attention), Manifold-Constrained Hyper-Connections, and Muon optimizer, V4-Flash-Max delivers reasoning performance comparable to V4-Pro when given a larger thinking budget while operating at a fraction of the parameter scale. It is pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation.
Radar de capacités
Classements
| Domaine | #Rang | Score | Source |
|---|---|---|---|
| Capacité agentique | 102 | 34.0 | LS |
| Classement codage | 79 | 87.0 | AA |
| Classement général | 169 | 57.0 | AA |
| Raisonnement mathématique | 50 | 80.0 | LB |
| Raisonnement | 51 | 71.0 | LB |
| Science | 88 | 76.0 | AA |
Scores de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Factuality
General
Language
Long Context
Math
Reasoning
Indices d'évaluation AA
(Artificial Analysis)Scores par catégorie LLM Stats
(LLM Stats (zeroeval))Tarification
Vitesse
Classement des Prix par Fournisseur
Classement des Prix par Fournisseur
39 fournisseurs
Comparer les prix entre différents fournisseurs API pour ce modèle.