Saltar al contenido principal

DeepSeek-V4-Flash-0423

DeepSeekDeepSeekOpen WeightMIT · Uso Comercial

Descripción

DeepSeek-V4-Flash-0423 is the preview release of DeepSeek-V4-Flash, a 284B-parameter MoE model with 13B activated parameters and a 1M-token context window, evaluated here at the default high reasoning effort. It shares the V4 series' hybrid attention architecture combining Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA) for dramatically improved long-context efficiency, Manifold-Constrained Hyper-Connections (mHC) for stable signal propagation, and the Muon optimizer for faster convergence. Pre-trained on more than 32T tokens and post-trained with a two-stage paradigm of domain-specific expert cultivation followed by on-policy distillation, V4-Flash offers reasoning capabilities that closely approach V4-Pro with faster responses and highly cost-effective pricing.

Fecha de lanzamiento
2026-04-23
Parámetros
284.0B
Longitud del contexto
Modalidades
text

Radar de capacidades

70
general
70
coding
80
reasoning
77
scienceest.
60
agents
0
multimodal

Science se estima a partir de las puntuaciones científicas de LLM Stats o del razonamiento cuando no hay benchmarks científicos dedicados.

Rankings

Dominio#PosiciónPuntuaciónFuente
Capacidad agéntica117
35.0
LS

Puntuaciones de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

Agents

MCP Atlas67.4%Aut.
Terminal-Bench 2.0Stanford × Laude Institute (2026)56.6%Aut.
BrowseCompOpenAI (2025)53.5%Aut.
SWE-Bench ProPrinceton NLP (2024)52.3%Aut.
Toolathlon43.5%Aut.

Biology

GPQANYU + Cohere + Anthropic (2023)87.4%Aut.

Code

LiveCodeBench88.4%Aut.
SWE-Bench Verified78.6%Aut.
SWE-bench Multilingual70.2%Aut.

Factuality

SimpleQA28.9%Aut.

Finance

MMLU-Pro86.4%Aut.

General

MRCR 1M76.9%Aut.
CSimpleQA73.2%Aut.
CorpusQA 1M59.3%Aut.

Math

CodeForces0.94 / 3000Aut.
HMMT Feb 2691.9%Aut.
IMO-AnswerBench85.1%Aut.
MathArena Apex72.1%Aut.
Humanity's Last Exam40.3%Aut.

Índices de evaluación AA

(Artificial Analysis)

No hay datos de evaluación AA disponibles

Puntuaciones por categoría LLM Stats

(LLM Stats (zeroeval))
Legal
90
Physics
90
Finance
90
Healthcare
90
Biology
90
Chemistry
90
Math
80
Language
80
Frontend Development
80
Long Context
70
Reasoning
70
General
70
Code
70
Tool Calling
60
Search
50
Agents
50
Vision
40
Factuality
30

Precios

Precio de entrada$0 / 1M tokens
Precio de salida$0 / 1M tokens
Precio mixto (3:1)$0 / 1M tokens

Velocidad

No hay datos de velocidad disponibles

Ranking de Precios por Proveedor

Ranking de Precios por Proveedor

3 proveedores

Más barato: DeepSeekMás caro: Novita
ProveedorEntradaSalida
1DeepSeekPRINCIPAL
$0
$0
2DeepInfra
$0
$0
3Novita
$0
$0

Comparar precios entre diferentes proveedores de API para este modelo.

Fuentes externas