DeepSeek-V2.5 (Dec '24)
DeepSeekDeepSeek오픈 웨이트deepseek
설명
DeepSeek-V2.5 is an upgraded version that combines DeepSeek-V2-Chat and DeepSeek-Coder-V2-Instruct, integrating general and coding abilities. It better aligns with human preferences and has been optimized in various aspects, including writing and instruction following.
출시일
2024-12-10
파라미터
236.0B
컨텍스트 길이
164K
모달리티
text
능력 레이더
28
general
60
coding
63
reasoning
42
science
62
agents
0
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Chat
MT-Bench
0.90 / 100자체 보고
General
MMLU
80.4%자체 보고
AlignBench
80.4%자체 보고
DS-FIM-Eval
78.3%자체 보고
Arena Hard
76.2%자체 보고
AlpacaEval 2.0
50.5%자체 보고
Math
GSM8k
95.1%자체 보고
MATH
74.7%자체 보고
Reasoning
HumanEvalOpenAI (2021)
89.0%자체 보고
BBH
84.3%자체 보고
HumanEval-Mul
73.8%자체 보고
Aider
72.2%자체 보고
DS-Arena-Code
63.1%자체 보고
LiveCodeBench(01-09)
41.8%자체 보고
SWE-Bench Verified
16.8%자체 보고
AA 평가 지수
(Artificial Analysis)Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))76.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))66.6
Gpqa(NYU + Cohere + Anthropic (2023))42.3
Intelligence Index(Artificial Analysis)6.6
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Roleplay90
Communication90
Chat80
Language80
Legal80
Math80
Finance80
Healthcare80
Reasoning70
General70
Creativity70
Writing70
Code60
Frontend Development20
가격
입력 가격무료
출력 가격무료
혼합 가격 (3:1)무료
속도
토큰/초0.0
첫 토큰 지연0.00s
첫 응답 지연0.00s
공급자 가격 순위
프로바이더 데이터가 없습니다