Does DeepSeek V4 Have Vision?
Blog post from OpenRouter
DeepSeek V4 is a family of models with differing input capabilities rather than a single vision-enabled model: DeepSeek V4.1 Flash and the experimental V4 Flash Vision Exp accept both text and images, while V4 Pro 0813, V4 Flash 0731, older 0423 versions, and the Flash “latest” alias are text-only. OpenRouter recommends V4.1 Flash for new image tasks because it supports native image understanding, a 1,048,576-token context window, tool calling, structured outputs, and lower listed pricing than the experimental alternative as of September 2026. Image requests can use public URLs or base64 data URLs, but sending images to text-only V4 models fails. To use a text-only V4 model, including V4 Pro, for image reasoning, users must first have a separate vision-capable model such as Qwen3.8 27B or Kimi K3 describe the image and then supply that description to V4, adding cost and removing direct pixel access. No listed V4 model supports video input, and users should verify current modalities, pricing, and aliases through OpenRouter model pages or its Models API before deploying a model.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| Real-time | 3 | 649 | 155 | 80 | -85% |
| Vector Search | 1 | 265 | 57 | 33 | -89% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.