Skip to main content

DeepSeek R1 0528 (May '25)

DeepSeekDeepSeekOpen WeightMIT · Commercial OK

Description

DeepSeek-R1-0528 is the May 28, 2025 version of DeepSeek's reasoning model. It features advanced thinking capabilities and serves as a benchmark comparison for newer models like DeepSeek-V3.1. This model excels in complex reasoning tasks, mathematical problem-solving, and code generation through its thinking mode approach.

Release Date
2025-05-28
Parameters
671.0B
Context Length
164K
Modalities
text

Capability Radar

35
general
77
coding
83
reasoning
61
science
10
agents
0
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability81
39.0
LS
Code Ranking256
57.0
AA
General Ranking326
41.0
AA
Science235
54.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

t2-bench80.2%SR
Toolathlon35.2%SR

Factuality

SimpleQA92.3%SR

General

Aider-Polyglot71.6%SR

Language

MMLU-Redux93.4%SR
MMLU-Pro85.0%SR

Math

AIME 202491.4%SR
AIME 202587.5%SR
HMMT 202579.4%SR
CodeForces0.64 / 3000SR

Reasoning

GPQANYU + Cohere + Anthropic (2023)81.0%SR
LiveCodeBench73.3%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)46.4%SR
SWE-Bench Verified44.6%SR
BrowseComp-zh35.7%SR
SWE-bench Multilingual30.5%SR
Humanity's Last Exam17.7%SR
BrowseCompOpenAI (2025)8.9%SR
Terminal-Bench5.7%SR

AA Evaluation Indices

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
98.3
Aime(MAA (Mathematical Association of America))
89.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
84.9
Gpqa(NYU + Cohere + Anthropic (2023))
81.3
Livecodebench(UC Berkeley + MIT + Cornell (2024))
77.0
Math Index(Artificial Analysis)
76.0
Aime 25(MAA (Mathematical Association of America))
76.0
Lcr(Artificial Analysis)
55.7
Ifbench(Google Research (2023))
39.6
Tau2(Sierra + U Toronto + Vector Institute (2025))
36.5
Terminalbench Hard(Stanford × Laude Institute (2026))
15.9
Hle(Center for AI Safety + Scale AI (2025))
15.8
Intelligence Index(Artificial Analysis)
13.1

LLM Stats Category Scores

(LLM Stats (zeroeval))
Language
90
Factuality
90
Legal
80
Physics
80
Finance
80
Healthcare
80
Biology
80
Chemistry
80
Math
70
Reasoning
60
General
60
Code
50
Frontend Development
40
Search
20
Vision
20
Agents
10

Pricing

Input Price$1.35 / 1M tokens
Output Price$3 / 1M tokens
Blended Price (3:1)$1.763 / 1M tokens
Cache Read Price$0.35 / 1M tokens

Speed

Tokens/sec0.0
Time to First Token0.00s
Time to Answer0.00s

Provider Price Ranking

Provider Price Ranking

11 providers

Cheapest: DeepInfraMost Expensive: Azure
ProviderInputOutput
1DeepInfraCheapest
$0
$0
2NanoGPT
$0.4
$1.7
3OpenRouter
$0.5
$2.15
4Alibaba (China)
$0.574
$2.294
5TensorX
$0.66
$2.6
6Jiekou.AI
$0.7
$2.5
7NovitaAI
$0.7
$2.5
8Kilo Gateway
$0.7
$2.5
9DeepSeekPRIMARY
$1.35
$3
10Azure Cognitive Services
$1.35
$5.4
11Azure
$1.35
$5.4

Compare pricing across different API providers for this model.

External Sources