DeepSeek R1 0528 (May '25)
DeepSeekDeepSeekओपन वेटMIT · व्यावसायिक उपयोग
विवरण
DeepSeek-R1-0528 is the May 28, 2025 version of DeepSeek's reasoning model. It features advanced thinking capabilities and serves as a benchmark comparison for newer models like DeepSeek-V3.1. This model excels in complex reasoning tasks, mathematical problem-solving, and code generation through its thinking mode approach.
रिलीज़ तिथि
2025-05-28
पैरामीटर
671.0B
संदर्भ लंबाई
164K
मोडैलिटीज़
text
क्षमता रडार
35
general
77
coding
83
reasoning
61
science
10
agents
0
multimodal
रैंकिंग
| डोमेन | #रैंक | स्कोर | स्रोत |
|---|---|---|---|
| एजेंटिक क्षमता | 81 | 39.0 | LS |
| कोडिंग रैंकिंग | 258 | 57.0 | AA |
| सामान्य रैंकिंग | 327 | 41.0 | AA |
| विज्ञान | 237 | 54.0 | AA |
बेंचमार्क स्कोर (LLM Stats)
(LLM Stats (zeroeval))Agents
t2-bench
80.2%स्वयं
Toolathlon
35.2%स्वयं
Factuality
SimpleQA
92.3%स्वयं
General
Aider-Polyglot
71.6%स्वयं
Language
MMLU-Redux
93.4%स्वयं
MMLU-Pro
85.0%स्वयं
Math
AIME 2024
91.4%स्वयं
AIME 2025
87.5%स्वयं
HMMT 2025
79.4%स्वयं
CodeForces
0.64 / 3000स्वयं
Reasoning
GPQANYU + Cohere + Anthropic (2023)
81.0%स्वयं
LiveCodeBench
73.3%स्वयं
Terminal-Bench 2.0Stanford × Laude Institute (2026)
46.4%स्वयं
SWE-Bench Verified
44.6%स्वयं
BrowseComp-zh
35.7%स्वयं
SWE-bench Multilingual
30.5%स्वयं
Humanity's Last Exam
17.7%स्वयं
BrowseCompOpenAI (2025)
8.9%स्वयं
Terminal-Bench
5.7%स्वयं
AA मूल्यांकन सूचकांक
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))98.3
Aime(MAA (Mathematical Association of America))89.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))84.9
Gpqa(NYU + Cohere + Anthropic (2023))81.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))77.0
Math Index(Artificial Analysis)76.0
Aime 25(MAA (Mathematical Association of America))76.0
Lcr(Artificial Analysis)55.7
Ifbench(Google Research (2023))39.6
Tau2(Sierra + U Toronto + Vector Institute (2025))36.5
Terminalbench Hard(Stanford × Laude Institute (2026))15.9
Hle(Center for AI Safety + Scale AI (2025))15.8
Intelligence Index(Artificial Analysis)13.1
LLM Stats श्रेणी स्कोर
(LLM Stats (zeroeval))Language90
Factuality90
Legal80
Physics80
Finance80
Healthcare80
Biology80
Chemistry80
Math70
Reasoning60
General60
Code50
Frontend Development40
Search20
Vision20
Agents10
मूल्य निर्धारण
इनपुट मूल्य$1.35 / 1M टोकन
आउटपुट मूल्य$3 / 1M टोकन
मिश्रित मूल्य (3:1)$1.763 / 1M टोकन
कैश पठन मूल्य$0.35 / 1M टोकन
गति
टोकन/सेकंड0.0
पहले टोकन में देरी0.00s
पहले उत्तर में देरी0.00s
प्रदाता मूल्य रैंकिंग
प्रदाता मूल्य रैंकिंग
11 प्रदाता
सबसे सस्ता: DeepInfraसबसे महंगा: Azure
प्रदाताइनपुटआउटपुट
1DeepInfraसबसे सस्ता
$0
$0
2NanoGPT
$0.4
$1.7
3OpenRouter
$0.5
$2.15
4Alibaba (China)
$0.574
$2.294
5TensorX
$0.66
$2.6
6Jiekou.AI
$0.7
$2.5
7NovitaAI
$0.7
$2.5
8Kilo Gateway
$0.7
$2.5
9DeepSeekप्राथमिक
$1.35
$3
10Azure Cognitive Services
$1.35
$5.4
11Azure
$1.35
$5.4
इस मॉडल के लिए विभिन्न API प्रदाताओं के मूल्य निर्धारण की तुलना करें।