Home / Companies / OpenRouter / Blog / Post Details
Content Deep Dive

Does DeepSeek V4 Have Vision?

Blog post from OpenRouter

Post Details
Company
Date Published
Author
OpenRouter
Word Count
1,798
Company Posts That Month
24
Language
English
Hacker News Points
-
Post removed?
No
Summary

DeepSeek V4 is a family of models with differing input capabilities rather than a single vision-enabled model: DeepSeek V4.1 Flash and the experimental V4 Flash Vision Exp accept both text and images, while V4 Pro 0813, V4 Flash 0731, older 0423 versions, and the Flash “latest” alias are text-only. OpenRouter recommends V4.1 Flash for new image tasks because it supports native image understanding, a 1,048,576-token context window, tool calling, structured outputs, and lower listed pricing than the experimental alternative as of September 2026. Image requests can use public URLs or base64 data URLs, but sending images to text-only V4 models fails. To use a text-only V4 model, including V4 Pro, for image reasoning, users must first have a separate vision-capable model such as Qwen3.8 27B or Kimi K3 describe the image and then supply that description to V4, adding cost and removing direct pixel access. No listed V4 model supports video input, and users should verify current modalities, pricing, and aliases through OpenRouter model pages or its Models API before deploying a model.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Real-time 3 649 155 80 -85%
Vector Search 1 265 57 33 -89%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.