DeepSeek: DeepSeek V4 Flash Vision Experimental
deepseek/deepseek-v4-flash-vision-exp
The deepseek-v4-flash-vision-exp model accepts images alongside text, so you can ask the model to describe pictures, read text from screenshots, analyze charts, and more. Supported image formats: JPEG, PNG, GIF, and WebP. The format is detected from the actual file content, not from the file name or the declared MIME type.
Model specifications
- Input
- text, image
- Output
- text
- Context
- 1,048,576 tokens
- Max output
- 384,000 tokens
- Input price
- $0.44 / 1M tokens
- Output price