메인 콘텐츠로 건너뛰기

DeepSeek V4.1 Flash (Reasoning, Max Effort)

DeepSeekDeepSeek오픈 웨이트MIT · 상업적 사용 가능

설명

DeepSeek-V4.1-Flash is an MIT-licensed multimodal Mixture-of-Experts model accepting images and text and generating text. It has 552B backbone parameters, 196B Engram conditional-memory parameters, and approximately 763B parameters in the released checkpoint. Its causal encoder-decoder architecture activates 8B parameters per token during prefill and 16B during decode. Trained on 45T multimodal tokens, it supports a 1M-token context, up to 384K output tokens on the DeepSeek API, and continuously adjustable reasoning effort from 1 to 100. CSA2 attention and FP4 KV caching reduce global KV cache storage to 890 bytes per token. The API model name is deepseek-flash. Catalog prices are peak rates per million tokens: $0.30 input, $0.006 cached input, and $1.20 output. Off-peak rates are $0.15, $0.003, and $0.60 respectively, effective September 10, 2026 at 04:00 UTC. Peak hours are 01:00–04:00 and 06:00–10:00 UTC on weekdays; all other times are off-peak (the launch pricing announcement also lists public holidays as off-peak).

출시일
2026-09-10
파라미터
763.2B
컨텍스트 길이
1.0M
모달리티
image, text

능력 레이더

39
general
52
coding
100
reasoning
47
science
50
agents
70
multimodal

랭킹

도메인#순위점수소스
에이전트형 역량41
49.0
LS
코딩 랭킹7
95.0
AA
종합 랭킹56
72.0
AA
수학 추론13
93.0
LB
멀티모달 랭킹4
73.0
LS
추론20
87.0
LB
과학95
72.0
AA

벤치마크 점수 (LLM Stats)

(LLM Stats (zeroeval))

Agents

SEC-bench Pro62.8%자체 보고
AutomationBench54.8%자체 보고
Agents' Last Exam31.8%자체 보고
Program Bench20.3%자체 보고
ExploitGym15.3%자체 보고

Code

CyberGym88.1%자체 보고
DeepSWE 1.174.2%자체 보고
NL2Repo64.0%자체 보고

Math

CodeForces3471.00 / 3000자체 보고
MathArena Apex65.6%자체 보고

Reasoning

GPQANYU + Cohere + Anthropic (2023)90.9%자체 보고
Terminal-Bench 2.190.6%자체 보고
Humanity's Last Exam (with tools)63.9%자체 보고
Humanity's Last Exam (no tools, text-only)39.1%자체 보고
Humanity's Last Exam36.8%자체 보고
Terminal-Bench 4.031.2%자체 보고
Terminal-Bench 3.030.0%자체 보고

Vision

BabyVision89.6%자체 보고
Chartography78.9%자체 보고
ZEROBench0.49 / 100자체 보고

AA 평가 지수

(Artificial Analysis)
Lcr(Artificial Analysis)
84.0
Scicode(UIUC + Argonne National Lab (2024))
51.9
Intelligence Index(Artificial Analysis)
39.5
Hle(Center for AI Safety + Scale AI (2025))
39.2

LLM Stats 카테고리 점수

(LLM Stats (zeroeval))
Math
100
Reasoning
100
General
100
Physics
90
Biology
90
Chemistry
90
Multimodal
70
Safety
60
Vision
60
Agents
50
Code
50
Tool Calling
50

가격

입력 가격$0.3 / 1M 토큰
출력 가격$1.2 / 1M 토큰
혼합 가격 (3:1)$0.525 / 1M 토큰
캐시 읽기 가격$0.003 / 1M 토큰

속도

토큰/초267.1
첫 토큰 지연0.90s
첫 응답 지연8.38s

공급자 가격 순위

공급자 가격 순위

13개 공급자

최저가: Fireworks최고가: Venice AI
공급자입력출력
1Fireworks최저가
$0
$0
2DeepSeek
$0
$0
3Novita
$0
$0
4DeepInfra
$0
$0
5NanoGPT
$0.15
$0.6
6OpenRouter
$0.15
$0.6
7Merge Gateway
$0.15
$0.6
8LLM Gateway
$0.15
$0.6
9CrossModel
$0.27
$1.08
10Kilo Gateway
$0.3
$1.2
11Vercel AI Gateway
$0.3
$1.2
12Ofox
$0.3
$1.2
13Venice AI
$0.375
$1.5

이 모델의 다양한 API 공급자 간 가격 비교.

외부 링크