DeepSeek V3 0324
DeepSeekDeepSeek開源權重MIT + Model License (Commercial use allowed)
描述
A powerful Mixture-of-Experts (MoE) language model with 671B total parameters (37B activated per token). Features Multi-head Latent Attention (MLA), auxiliary-loss-free load balancing, and multi-token prediction training. Pre-trained on 14.8T tokens with strong performance in reasoning, math, and code tasks.
發布日期
2025-03-25
參數規模
671.0B
上下文長度
164K
支援模態
text
能力雷達圖
31
general
30
coding
54
reasoning
44
science
47
agents
0
multimodal
排行榜排名
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Language
MMLU-Pro
81.2%自報
Math
MATH-500
94.0%自報
AIME 2024
59.4%自報
Reasoning
GPQANYU + Cohere + Anthropic (2023)
68.4%自報
LiveCodeBench
49.2%自報
AA 評測指數
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))94.2
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))81.9
Gpqa(NYU + Cohere + Anthropic (2023))65.5
Aime(MAA (Mathematical Association of America))52.0
Tau2(Sierra + U Toronto + Vector Institute (2025))47.1
Ifbench(Google Research (2023))41.0
Math Index(Artificial Analysis)41.0
Aime 25(MAA (Mathematical Association of America))41.0
Lcr(Artificial Analysis)40.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))40.5
Scicode(UIUC + Argonne National Lab (2024))39.0
Coding Index(Artificial Analysis)21.2
Terminalbench Hard(Stanford × Laude Institute (2026))15.2
Terminalbench V2 113.9
Intelligence Index(Artificial Analysis)9.7
Tau Banking4.7
Hle(Center for AI Safety + Scale AI (2025))4.7
Terminalbench V4 00.0
LLM Stats 分類評分
(LLM Stats (zeroeval))Language80
Legal80
Math80
Finance80
Healthcare80
Physics70
Reasoning70
General70
Biology70
Chemistry70
Code50
定價
輸入價格$0.24 / 1M tokens
輸出價格$0.9 / 1M tokens
混合價格(3:1)$0.405 / 1M tokens
快取讀取價格$0.11 / 1M tokens
速度
Tokens/秒0.0
首Token延遲0.00s
首回答延遲0.00s
供應商價格排行
供應商價格排行
5 個供應商
最便宜: DeepInfra最貴: Kilo Gateway
供應商輸入輸出
1DeepInfra最便宜
$0
$0
2NanoGPT
$0.2
$0.77
3DeepSeek主要
$0.24
$0.9
4OpenRouter
$0.29
$1.14
5Kilo Gateway
$0.29
$1.14
比較該模型在不同 API 供應商之間的定價。