DeepSeek-V4-Flash-Vision-Exp
描述
DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal vision-understanding model on the DeepSeek API (`model='deepseek-v4-flash-vision-exp'`). It accepts text and image inputs and produces text. DeepSeek reports pure-text agent, reasoning, and world-knowledge performance on par with official DeepSeek-V4-Flash, with a large jump on agent benchmarks that require visual understanding. Thinking mode is default and supports low / high / max effort; the model supports JSON output, tool calls, the Responses API, and the Anthropic API. FIM completion is not supported. Context length is 1M tokens with up to 384K max output. Same peak/off-peak token pricing as V4-Flash; images are converted to input tokens (upper bound 384 tokens per image after resize).
能力雷達圖
缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。
排行榜排名
| 領域 | #排名 | 分數 | 來源 |
|---|---|---|---|
| 多模態榜 | 21 | 62.0 | LS |
基準測試分數 (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Reasoning
Vision
AA 評測指數
(Artificial Analysis)暫無 AA 評測資料
LLM Stats 分類評分
(LLM Stats (zeroeval))定價
速度
暫無速度資料
供應商價格排行
供應商價格排行
8 個供應商
比較該模型在不同 API 供應商之間的定價。