DeepSeek V3 0324
DeepSeekDeepSeek开源权重MIT + Model License (Commercial use allowed)
描述
A powerful Mixture-of-Experts (MoE) language model with 671B total parameters (37B activated per token). Features Multi-head Latent Attention (MLA), auxiliary-loss-free load balancing, and multi-token prediction training. Pre-trained on 14.8T tokens with strong performance in reasoning, math, and code tasks.
发布日期
2025-03-25
参数规模
671.0B
上下文长度
164K
支持模态
text
能力雷达图
31
general
30
coding
54
reasoning
44
science
47
agents
0
multimodal
排行榜排名
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Language
MMLU-Pro
81.2%自报
Math
MATH-500
94.0%自报
AIME 2024
59.4%自报
Reasoning
GPQANYU + Cohere + Anthropic (2023)
68.4%自报
LiveCodeBench
49.2%自报
AA 评测指数
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))94.2
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))81.9
Gpqa(NYU + Cohere + Anthropic (2023))65.5
Aime(MAA (Mathematical Association of America))52.0
Tau2(Sierra + U Toronto + Vector Institute (2025))47.1
Ifbench(Google Research (2023))41.0
Math Index(Artificial Analysis)41.0
Aime 25(MAA (Mathematical Association of America))41.0
Lcr(Artificial Analysis)40.7
Livecodebench(UC Berkeley + MIT + Cornell (2024))40.5
Scicode(UIUC + Argonne National Lab (2024))39.0
Coding Index(Artificial Analysis)21.2
Terminalbench Hard(Stanford × Laude Institute (2026))15.2
Terminalbench V2 113.9
Intelligence Index(Artificial Analysis)9.7
Tau Banking4.7
Hle(Center for AI Safety + Scale AI (2025))4.7
Terminalbench V4 00.0
LLM Stats 分类评分
(LLM Stats (zeroeval))Language80
Legal80
Math80
Finance80
Healthcare80
Physics70
Reasoning70
General70
Biology70
Chemistry70
Code50
定价
输入价格$0.24 / 1M tokens
输出价格$0.9 / 1M tokens
混合价格(3:1)$0.405 / 1M tokens
缓存读取价格$0.11 / 1M tokens
速度
Tokens/秒0.0
首Token延迟0.00s
首回答延迟0.00s
供应商价格排行
供应商价格排行
5 个供应商
最便宜: DeepInfra最贵: Kilo Gateway
供应商输入输出
1DeepInfra最便宜
$0
$0
2NanoGPT
$0.2
$0.77
3DeepSeek主要
$0.24
$0.9
4OpenRouter
$0.29
$1.14
5Kilo Gateway
$0.29
$1.14
比较该模型在不同 API 供应商之间的定价。