DeepSeek-V4-Flash-Vision-Exp
Описание
DeepSeek-V4-Flash-Vision-Exp is an experimental multimodal vision-understanding model on the DeepSeek API (`model='deepseek-v4-flash-vision-exp'`). It accepts text and image inputs and produces text. DeepSeek reports pure-text agent, reasoning, and world-knowledge performance on par with official DeepSeek-V4-Flash, with a large jump on agent benchmarks that require visual understanding. Thinking mode is default and supports low / high / max effort; the model supports JSON output, tool calls, the Responses API, and the Anthropic API. FIM completion is not supported. Context length is 1M tokens with up to 384K max output. Same peak/off-peak token pricing as V4-Flash; images are converted to input tokens (upper bound 384 tokens per image after resize).
Радар способностей
При отсутствии специализированных научных бенчмарков Science оценивается по научным категориям LLM Stats или на основе рассуждений.
Рейтинги
| Домен | #Место | Оценка | Источник |
|---|---|---|---|
| Мультимодальный рейтинг | 21 | 62.0 | LS |
Оценки бенчмарков (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Reasoning
Vision
Индексы оценки AA
(Artificial Analysis)Нет данных AA оценки
Оценки категорий LLM Stats
(LLM Stats (zeroeval))Цены
Скорость
Нет данных о скорости
Рейтинг цен провайдеров
Рейтинг цен провайдеров
8 провайдеров
Сравнение цен разных API-провайдеров для этой модели.