Questions on DeepSeek V4 Pro Image Input

Porn197617_ · reddit · 2026-07-13

While building an AI writing agent workflow, the author switched the primary model from GPT-5.5 to DeepSeek V4 Pro and found that its API seemingly cannot directly read images.

Their usual workflow involves feeding design drafts, UI screenshots, or error screenshots to the agent so it understands the goal before proceeding. Without image capabilities, many layout and visual details must be described in text, significantly dropping efficiency. The author wants to confirm:

Original post →

More from coding & agent

coding & agent channel →