DeepSeek VL2 Tiny
DeepSeekDeepSeekOpen Weightdeepseek
Description
An advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly improves upon its predecessor, DeepSeek-VL. DeepSeek-VL2 demonstrates superior capabilities across various tasks, including but not limited to visual question answering, optical character recognition, document/table/chart understanding, and visual grounding.
Release Date
2024-12-13
Parameters
3.0B
Context Length
—
Modalities
—
Capability Radar
60
general
0
coding
50
reasoning
34
scienceest.
35
agents
80
multimodal
Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Multimodal Ranking | 84 | 23.0 | LS |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))General
MMT-Bench
53.2%SR
MMStar
45.9%SR
MMMU
40.7%SR
Image To Text
DocVQADocVQA (2020)
88.9%SR
OCRBench
80.9%SR
TextVQA
80.7%SR
Math
MathVista
53.6%SR
Multimodal
ChartQAMasry et al. (2022)
81.0%SR
AI2D
71.6%SR
MMBench
69.2%SR
MMBench-V1.1
68.3%SR
InfoVQA
66.1%SR
MME
19.1%SR
Spatial Reasoning
RealWorldQA
64.2%SR
AA Evaluation Indices
(Artificial Analysis)No AA evaluation data available
LLM Stats Category Scores
(LLM Stats (zeroeval))Image To Text80
Multimodal60
Reasoning60
Spatial Reasoning60
General60
Vision60
Math50
Healthcare40
Pricing
No pricing data available
Speed
No speed data available
Provider Price Ranking
No provider data available