DeepSeek V3.2 (Non-reasoning)
DeepSeekDeepSeekオープンウエイトMIT · 商用利用可
説明
DeepSeek-V3.2 is a 685B-parameter MoE model that harmonizes high computational efficiency with superior reasoning and agent performance. It introduces DeepSeek Sparse Attention (DSA) for efficient long-context processing, a scalable reinforcement learning post-training framework, and large-scale agentic task synthesis covering 1,800+ environments. V3.2 achieves GPT-5-level performance across reasoning, coding, and agentic benchmarks, with gold-medal results from its Speciale variant on IMO, IOI, ICPC World Finals, and CMO 2025.
リリース日
2025-12-01
パラメータ
685.0B
コンテキスト長
164K
モダリティ
text
能力レーダー
41
general
55
coding
62
reasoning
50
science
50
agents
0
multimodal
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 135 | 31.0 | LS |
| コーディングランキング | 185 | 55.0 | AA |
| 総合ランキング | 154 | 61.0 | AA |
| 科学 | 208 | 53.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
80.3%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)
46.4%自己申告
MCP-Universe
45.9%自己申告
BrowseCompOpenAI (2025)
40.1%自己申告
MCP-Mark
38.0%自己申告
Terminal-Bench
37.7%自己申告
Toolathlon
35.2%自己申告
Biology
GPQANYU + Cohere + Anthropic (2023)
79.9%自己申告
Code
Aider-Polyglot
74.5%自己申告
LiveCodeBench
74.1%自己申告
SWE-Bench Verified
67.8%自己申告
SWE-bench Multilingual
57.9%自己申告
Factuality
SimpleQA
97.1%自己申告
Finance
MMLU-Pro
85.0%自己申告
Math
AIME 2025
89.3%自己申告
HMMT 2025
83.6%自己申告
IMO-AnswerBench
78.3%自己申告
CodeForces
0.71 / 3000自己申告
Humanity's Last Exam
19.8%自己申告
Reasoning
BrowseComp-zh
47.9%自己申告
AA評価指数
(Artificial Analysis)Math Index(Artificial Analysis)59.0
Intelligence Index(Artificial Analysis)25.1
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))0.8
Tau2(Sierra + U Toronto + Vector Institute (2025))0.8
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))0.6
Aime 25(MAA (Mathematical Association of America))0.6
Ifbench(Google Research (2023))0.5
Lcr(Artificial Analysis)0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Terminalbench Hard(Stanford × Laude Institute (2026))0.3
Hle(Center for AI Safety + Scale AI (2025))0.1
LLM Statsカテゴリスコア
(LLM Stats (zeroeval))Language80
Legal80
Math80
Physics80
Finance80
Healthcare80
Biology80
Chemistry80
Reasoning70
Frontend Development70
General70
Code70
Search60
Agents50
Tool Calling50
Vision40
価格設定
入力価格$0.28 / 1Mトークン
出力価格$0.42 / 1Mトークン
混合価格(3:1)$0.315 / 1Mトークン
キャッシュ読み取り価格$0.1345 / 1Mトークン
速度
トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s
プロバイダー価格ランキング
プロバイダー価格ランキング
2 プロバイダー
最安: DeepSeek最高: EmpirioLabs AI
プロバイダー入力出力
1DeepSeekプライマリ
$0.28
$0.42
2EmpirioLabs AI
$0.57
$1.71
このモデルの異なるAPIプロバイダー間の価格を比較。