DeepSeek adds vision API support via deepseek-v4-flash-vision-exp model
Original titleMultimodal API support 🔌
AISummary
DeepSeek's API now accepts multimodal input through the model deepseek-v4-flash-vision-exp, supporting mixed text and image requests. Each image is billed at up to 384 tokens at V4-Flash pricing, and it works with Chat Completions, Messages, and Responses endpoints. Images can be supplied as base64, external URLs, or via the Files API.
Source: DeepSeek · x.comPublished · added here