DeepSeek-V4-Flash-Vision-Exp
描述
DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal vision-understanding model on the DeepSeek API (`model='deepseek-v4-flash-vision-exp'`). It accepts text and image inputs and produces text. DeepSeek reports pure-text agent, reasoning, and world-knowledge performance on par with official DeepSeek-V4-Flash, with a large jump on agent benchmarks that require visual understanding. Thinking mode is default and supports low / high / max effort; the model supports JSON output, tool calls, the Responses API, and the Anthropic API. FIM completion is not supported. Context length is 1M tokens with up to 384K max output. Same peak/off-peak token pricing as V4-Flash; images are converted to input tokens (upper bound 384 tokens per image after resize).
能力雷达图
缺少专门科学评测时,Science 由 LLM Stats 科学得分或推理能力估算。
排行榜排名
| 领域 | #排名 | 分数 | 来源 |
|---|---|---|---|
| 多模态榜 | 21 | 62.0 | LS |
基准测试分数 (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Reasoning
Vision
AA 评测指数
(Artificial Analysis)暂无 AA 评测数据
LLM Stats 分类评分
(LLM Stats (zeroeval))定价
速度
暂无速度数据
供应商价格排行
供应商价格排行
8 个供应商
比较该模型在不同 API 供应商之间的定价。