メインコンテンツへスキップ

DeepSeek V3.2 (Non-reasoning)

DeepSeekDeepSeekオープンウエイトMIT · 商用利用可

説明

DeepSeek-V3.2 is a 685B-parameter MoE model that harmonizes high computational efficiency with superior reasoning and agent performance. It introduces DeepSeek Sparse Attention (DSA) for efficient long-context processing, a scalable reinforcement learning post-training framework, and large-scale agentic task synthesis covering 1,800+ environments. V3.2 achieves GPT-5-level performance across reasoning, coding, and agentic benchmarks, with gold-medal results from its Speciale variant on IMO, IOI, ICPC World Finals, and CMO 2025.

リリース日
2025-12-01
パラメータ
685.0B
コンテキスト長
164K
モダリティ
text

能力レーダー

36
general
59
coding
62
reasoning
55
science
50
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
エージェント能力94
36.0
LS
コーディングランキング267
55.0
AA
総合ランキング195
55.0
AA
科学303
46.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench80.3%自己申告
MCP-Universe45.9%自己申告
MCP-Mark38.0%自己申告
Toolathlon35.2%自己申告

Factuality

SimpleQA97.1%自己申告

General

Aider-Polyglot74.5%自己申告

Language

MMLU-Pro85.0%自己申告

Math

AIME 202589.3%自己申告
HMMT 202583.6%自己申告
IMO-AnswerBench78.3%自己申告
CodeForces0.71 / 3000自己申告

Reasoning

GPQANYU + Cohere + Anthropic (2023)79.9%自己申告
LiveCodeBench74.1%自己申告
SWE-Bench Verified67.8%自己申告
SWE-bench Multilingual57.9%自己申告
BrowseComp-zh47.9%自己申告
Terminal-Bench 2.0Stanford × Laude Institute (2026)46.4%自己申告
BrowseCompOpenAI (2025)40.1%自己申告
Terminal-Bench37.7%自己申告
Humanity's Last Exam19.8%自己申告

AA評価指数

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
83.7
Tau2(Sierra + U Toronto + Vector Institute (2025))
78.9
Gpqa(NYU + Cohere + Anthropic (2023))
75.1
Livecodebench(UC Berkeley + MIT + Cornell (2024))
59.3
Math Index(Artificial Analysis)
59.0
Aime 25(MAA (Mathematical Association of America))
59.0
Ifbench(Google Research (2023))
49.0
Lcr(Artificial Analysis)
45.7
Terminalbench Hard(Stanford × Laude Institute (2026))
32.6
Intelligence Index(Artificial Analysis)
16.0
Hle(Center for AI Safety + Scale AI (2025))
11.2

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Language
80
Legal
80
Math
80
Physics
80
Finance
80
Healthcare
80
Biology
80
Chemistry
80
Reasoning
70
Frontend Development
70
General
70
Code
70
Search
60
Agents
50
Tool Calling
50
Vision
40

価格設定

入力価格$0.28 / 1Mトークン
出力価格$0.42 / 1Mトークン
混合価格(3:1)$0.315 / 1Mトークン
キャッシュ読み取り価格$0.028 / 1Mトークン

速度

トークン/秒0.0
初トークン遅延0.00s
初回答遅延0.00s

プロバイダー価格ランキング

プロバイダー価格ランキング

13 プロバイダー

最安: DeepInfra最高: Vercel AI Gateway
プロバイダー入力出力
1DeepInfra最安
$0
$0
2TokenGo
$0.2174
$0.326
3NovitaAI
$0.269
$0.4
4DeepSeekプライマリ
$0.28
$0.42
5NanoGPT
$0.28
$0.42
6OpenRouter
$0.28
$0.42
7ZenMux
$0.28
$0.43
8Kilo Gateway
$0.28
$0.42
9Merge Gateway
$0.28
$0.45
10Ofox
$0.29
$0.43
11TensorX
$0.3
$0.5
12EmpirioLabs AI
$0.57
$1.71
13Vercel AI Gateway
$0.62
$1.85

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク