跳轉到主要內容

DeepSeek V4.1 Flash (Reasoning, Max Effort)

DeepSeekDeepSeek開源權重MIT · 商用許可

描述

DeepSeek-V4.1-Flash is an MIT-licensed multimodal Mixture-of-Experts model accepting images and text and generating text. It has 552B backbone parameters, 196B Engram conditional-memory parameters, and approximately 763B parameters in the released checkpoint. Its causal encoder-decoder architecture activates 8B parameters per token during prefill and 16B during decode. Trained on 45T multimodal tokens, it supports a 1M-token context, up to 384K output tokens on the DeepSeek API, and continuously adjustable reasoning effort from 1 to 100. CSA2 attention and FP4 KV caching reduce global KV cache storage to 890 bytes per token. The API model name is deepseek-flash. Catalog prices are peak rates per million tokens: $0.30 input, $0.006 cached input, and $1.20 output. Off-peak rates are $0.15, $0.003, and $0.60 respectively, effective September 10, 2026 at 04:00 UTC. Peak hours are 01:00–04:00 and 06:00–10:00 UTC on weekdays; all other times are off-peak (the launch pricing announcement also lists public holidays as off-peak).

發布日期
2026-09-10
參數規模
763.2B
上下文長度
1.0M
支援模態
image, text

能力雷達圖

39
general
52
coding
100
reasoning
47
science
50
agents
70
multimodal

排行榜排名

領域#排名分數來源
智慧體能力模型榜41
49.0
LS
程式碼能力榜7
95.0
AA
通用能力榜56
72.0
AA
數學推理13
93.0
LB
多模態榜4
73.0
LS
推理能力20
87.0
LB
科學能力95
72.0
AA

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Agents

SEC-bench Pro62.8%自報
AutomationBench54.8%自報
Agents' Last Exam31.8%自報
Program Bench20.3%自報
ExploitGym15.3%自報

Code

CyberGym88.1%自報
DeepSWE 1.174.2%自報
NL2Repo64.0%自報

Math

CodeForces3471.00 / 3000自報
MathArena Apex65.6%自報

Reasoning

GPQANYU + Cohere + Anthropic (2023)90.9%自報
Terminal-Bench 2.190.6%自報
Humanity's Last Exam (with tools)63.9%自報
Humanity's Last Exam (no tools, text-only)39.1%自報
Humanity's Last Exam36.8%自報
Terminal-Bench 4.031.2%自報
Terminal-Bench 3.030.0%自報

Vision

BabyVision89.6%自報
Chartography78.9%自報
ZEROBench0.49 / 100自報

AA 評測指數

(Artificial Analysis)
Lcr(Artificial Analysis)
84.0
Scicode(UIUC + Argonne National Lab (2024))
51.9
Intelligence Index(Artificial Analysis)
39.5
Hle(Center for AI Safety + Scale AI (2025))
39.2

LLM Stats 分類評分

(LLM Stats (zeroeval))
Math
100
Reasoning
100
General
100
Physics
90
Biology
90
Chemistry
90
Multimodal
70
Safety
60
Vision
60
Agents
50
Code
50
Tool Calling
50

定價

輸入價格$0.3 / 1M tokens
輸出價格$1.2 / 1M tokens
混合價格(3:1)$0.525 / 1M tokens
快取讀取價格$0.003 / 1M tokens

速度

Tokens/秒267.1
首Token延遲0.90s
首回答延遲8.38s

供應商價格排行

供應商價格排行

13 個供應商

最便宜: Fireworks最貴: Venice AI
供應商輸入輸出
1Fireworks最便宜
$0
$0
2DeepSeek
$0
$0
3Novita
$0
$0
4DeepInfra
$0
$0
5NanoGPT
$0.15
$0.6
6OpenRouter
$0.15
$0.6
7Merge Gateway
$0.15
$0.6
8LLM Gateway
$0.15
$0.6
9CrossModel
$0.27
$1.08
10Kilo Gateway
$0.3
$1.2
11Vercel AI Gateway
$0.3
$1.2
12Ofox
$0.3
$1.2
13Venice AI
$0.375
$1.5

比較該模型在不同 API 供應商之間的定價。

外部連結