DeepSeek visual model officially launched API: the price is exactly the same as V4 Flash
DeepSeek officially launches the new visual model deepseek-v4-flash-vision-exp. The official API documentation has listed it alongside V4 Flash and V4 Pro, allowing developers to directly input images through the DeepSeek API. The model supports a context of 1 million tokens, with a maximum output of 384,000 tokens, and also supports JSON Output, Tool Calls, Responses API, and Anthropic API.
The pricing is directly aligned with V4 Flash. For every million tokens input, when there is a cache miss, the peak period is 3 yuan, and the idle period is 1.5 yuan; cache hits only cost 0.1 yuan and 0.05 yuan respectively. For every million tokens output, the peak period is 9 yuan, and the idle period is 4.5 yuan. Compared to V4 Pro, the same tier pricing is only one-third.
Images are not charged separately per image. DeepSeek will convert images into tokens based on their size and bill them together with text tokens. The peak periods are from 9:00 to 12:00 and 14:00 to 18:00 Beijing time, with prices halved during other times.






