メインコンテンツへスキップ

DeepSeek V4.1 Flash (Reasoning, Max Effort)

DeepSeekDeepSeekオープンウエイトMIT · 商用利用可

説明

DeepSeek-V4.1-Flash is an MIT-licensed multimodal Mixture-of-Experts model accepting images and text and generating text. It has 552B backbone parameters, 196B Engram conditional-memory parameters, and approximately 763B parameters in the released checkpoint. Its causal encoder-decoder architecture activates 8B parameters per token during prefill and 16B during decode. Trained on 45T multimodal tokens, it supports a 1M-token context, up to 384K output tokens on the DeepSeek API, and continuously adjustable reasoning effort from 1 to 100. CSA2 attention and FP4 KV caching reduce global KV cache storage to 890 bytes per token. The API model name is deepseek-flash. Catalog prices are peak rates per million tokens: $0.30 input, $0.006 cached input, and $1.20 output. Off-peak rates are $0.15, $0.003, and $0.60 respectively, effective September 10, 2026 at 04:00 UTC. Peak hours are 01:00–04:00 and 06:00–10:00 UTC on weekdays; all other times are off-peak (the launch pricing announcement also lists public holidays as off-peak).

リリース日
2026-09-10
パラメータ
763.2B
コンテキスト長
1.0M
モダリティ
image, text

能力レーダー

39
general
52
coding
100
reasoning
47
science
50
agents
70
multimodal

ランキング

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

SEC-bench Pro62.8%自己申告
AutomationBench54.8%自己申告
Agents' Last Exam31.8%自己申告
Program Bench20.3%自己申告
ExploitGym15.3%自己申告

Code

CyberGym88.1%自己申告
DeepSWE 1.174.2%自己申告
NL2Repo64.0%自己申告

Math

CodeForces3471.00 / 3000自己申告
MathArena Apex65.6%自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)90.9%自己申告
Terminal-Bench 2.190.6%自己申告
Humanity's Last Exam (with tools)63.9%自己申告
Humanity's Last Exam (no tools, text-only)39.1%自己申告
Humanity's Last Exam36.8%自己申告
Terminal-Bench 4.031.2%自己申告
Terminal-Bench 3.030.0%自己申告

Vision

BabyVision89.6%自己申告
Chartography78.9%自己申告
ZEROBench0.49 / 100自己申告

AA評価指数

(Artificial Analysis)
Lcr(Artificial Analysis)
84.0
Scicode(UIUC + Argonne National Lab (2024))
51.9
Intelligence Index(Artificial Analysis)
39.5
Hle(Center for AI Safety + Scale AI (2025))
39.2

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Math
100
Reasoning
100
General
100
Physics
90
Biology
90
Chemistry
90
Multimodal
70
Safety
60
Vision
60
Agents
50
Code
50
Tool Calling
50

価格設定

入力価格$0.3 / 1Mトークン
出力価格$1.2 / 1Mトークン
混合価格(3:1)$0.525 / 1Mトークン
キャッシュ読み取り価格$0.003 / 1Mトークン

速度

トークン/秒267.1
初トークン遅延0.90s
初回答遅延8.38s

プロバイダー価格ランキング

プロバイダー価格ランキング

13 プロバイダー

最安: Fireworks最高: Venice AI
プロバイダー入力出力
1Fireworks最安
$0
$0
2DeepSeek
$0
$0
3Novita
$0
$0
4DeepInfra
$0
$0
5NanoGPT
$0.15
$0.6
6OpenRouter
$0.15
$0.6
7Merge Gateway
$0.15
$0.6
8LLM Gateway
$0.15
$0.6
9CrossModel
$0.27
$1.08
10Kilo Gateway
$0.3
$1.2
11Vercel AI Gateway
$0.3
$1.2
12Ofox
$0.3
$1.2
13Venice AI
$0.375
$1.5

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク