メインコンテンツへスキップ

DeepSeek VL2

DeepSeekDeepSeekオープンウエイトdeepseek

説明

An advanced series of large Mixture-of-Experts (MoE) Vision-Language Models that significantly improves upon its predecessor, DeepSeek-VL. DeepSeek-VL2 demonstrates superior capabilities across various tasks, including but not limited to visual question answering, optical character recognition, document/table/chart understanding, and visual grounding.

リリース日
2024-12-13
パラメータ
27.0B
コンテキスト長
モダリティ
image, text

能力レーダー

70
general
0
coding
60
reasoning
43
science推定
42
agents
90
multimodal

専用の科学ベンチマークがない場合、Science は LLM Stats の科学スコアまたは推論能力から推定します。

ランキング

ドメイン#順位スコアソース
マルチモーダルランキング78
25.0
LS

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

General

MMT-Bench63.6%自己申告
MMStar61.3%自己申告
MMMU51.1%自己申告

Image To Text

DocVQADocVQA (2020)93.3%自己申告
TextVQA84.2%自己申告
OCRBench81.1%自己申告

Math

MathVista62.8%自己申告

Multimodal

ChartQAMasry et al. (2022)86.0%自己申告
AI2D81.4%自己申告
MMBench79.6%自己申告
MMBench-V1.179.2%自己申告
InfoVQA78.1%自己申告
MME22.5%自己申告

Spatial Reasoning

RealWorldQA68.4%自己申告

AA評価指数

(Artificial Analysis)

AA評価データがありません

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Image To Text
90
Multimodal
70
Reasoning
70
Spatial Reasoning
70
General
70
Vision
70
Math
60
Healthcare
50

価格設定

価格データがありません

速度

速度データがありません

プロバイダー価格ランキング

プロバイダーデータがありません

外部リンク