DeepSeek VL2
DeepSeekDeepSeek開源權重deepseek
描述
An advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly improves upon its predecessor, DeepSeek-VL. DeepSeek-VL2 demonstrates superior capabilities across various tasks, including but not limited to visual question answering, optical character recognition, document/table/chart understanding, and visual grounding.
發布日期
2024-12-13
參數規模
27.0B
上下文長度
—
支援模態
image, text
能力雷達圖
70
general
0
coding
60
reasoning
43
science估算
42
agents
90
multimodal
缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。
排行榜排名
| 領域 | #排名 | 分數 | 來源 |
|---|---|---|---|
| 多模態榜 | 79 | 28.0 | LS |
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))General
MMT-Bench
63.6%自報
MMStar
61.3%自報
MMMU
51.1%自報
Image To Text
DocVQADocVQA (2020)
93.3%自報
TextVQA
84.2%自報
OCRBench
81.1%自報
Math
MathVista
62.8%自報
Multimodal
ChartQAMasry et al. (2022)
86.0%自報
AI2D
81.4%自報
MMBench
79.6%自報
MMBench-V1.1
79.2%自報
InfoVQA
78.1%自報
MME
22.5%自報
Spatial Reasoning
RealWorldQA
68.4%自報
AA 評測指數
(Artificial Analysis)暫無 AA 評測資料
LLM Stats 分類評分
(LLM Stats (zeroeval))Image To Text90
Multimodal70
Reasoning70
Spatial Reasoning70
General70
Vision70
Math60
Healthcare50
定價
暫無定價資料
速度
暫無速度資料
供應商價格排行
暫無提供商資料