メインコンテンツへスキップ

DeepSeek V3.2 (Non-reasoning)

DeepSeekDeepSeekオープンウエイトMIT · 商用利用可

説明

DeepSeek-V3.2 is a 685B-parameter MoE model that harmonizes high computational efficiency with superior reasoning and agent performance. It introduces DeepSeek Sparse Attention (DSA) for efficient long-context processing, a scalable reinforcement learning post-training framework, and large-scale agentic task synthesis covering 1,800+ environments. V3.2 achieves GPT-5-level performance across reasoning, coding, and agentic benchmarks, with gold-medal results from its Speciale variant on IMO, IOI, ICPC World Finals, and CMO 2025.

リリース日
2025-12-01
パラメータ
685.0B
コンテキスト長
164K
モダリティ
text

能力レーダー

41
general
55
coding
62
reasoning
50
science
50
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
エージェント能力135
31.0
LS
コーディングランキング185
55.0
AA
総合ランキング154
61.0
AA
科学208
53.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench80.3%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)46.4%自己申告
MCP-Universe45.9%自己申告
BrowseCompOpenAI (2025)40.1%自己申告
MCP-Mark38.0%自己申告
Terminal-Bench37.7%自己申告
Toolathlon35.2%自己申告

Biology

GPQANYU + Cohere + Anthropic (2023)79.9%自己申告

Code

Aider-Polyglot74.5%自己申告
LiveCodeBench74.1%自己申告
SWE-Bench Verified67.8%自己申告
SWE-bench Multilingual57.9%自己申告

Factuality

SimpleQA97.1%自己申告

Finance

MMLU-Pro85.0%自己申告

Math

AIME 202589.3%自己申告
HMMT 202583.6%自己申告
IMO-AnswerBench78.3%自己申告
CodeForces0.71 / 3000自己申告
Humanity's Last Exam19.8%自己申告

Reasoning

BrowseComp-zh47.9%自己申告

AA評価指数

(Artificial Analysis)
Math Index(Artificial Analysis)
59.0
Intelligence Index(Artificial Analysis)
25.1
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
0.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
0.8
Gpqa(NYU + Cohere + Anthropic (2023))
0.8
Livecodebench(UC Berkeley + MIT + Cornell (2024))
0.6
Aime 25(MAA (Mathematical Association of America))
0.6
Ifbench(Google Research (2023))
0.5
Lcr(Artificial Analysis)
0.4
Scicode(UIUC + Argonne National Lab (2024))
0.4
Terminalbench Hard(Stanford × Laude Institute (2026))
0.3
Hle(Center for AI Safety + Scale AI (2025))
0.1

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Language
80
Legal
80
Math
80
Physics
80
Finance
80
Healthcare
80
Biology
80
Chemistry
80
Reasoning
70
Frontend Development
70
General
70
Code
70
Search
60
Agents
50
Tool Calling
50
Vision
40

価格設定

入力価格$0.28 / 1Mトークン
出力価格$0.42 / 1Mトークン
混合価格(3:1)$0.315 / 1Mトークン
キャッシュ読み取り価格$0.1345 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

2 プロバイダー

最安: DeepSeek最高: EmpirioLabs AI
プロバイダー入力出力
1DeepSeekプライマリ
$0.28
$0.42
2EmpirioLabs AI
$0.57
$1.71

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク