DeepSeekHarness update adds multimodal support and Codex sub-agents

智东西 · wechat · 2026-08-20

DeepSeekHarness version v0.1.0-rc.8 introduces significant multimodal capabilities, supporting native image requests and mixed text-image inputs for commands like /goal and /plan. The update integrates ClaudeCode and Codex into the sub-agent ecosystem, supporting non-interactive modes and multiple named instances for Codex. For image processing, the system allows visual models to input images directly, while text-only models can use tools like OCR, color statistics, and pixel scanning to interpret images indirectly. It also fixes Windows terminal issues, image request failures, and streaming generation bugs, alongside SQLite backend performance improvements.

Related event: DeepSeekHarness RC.8 Adds Multimodal Support and Subagents(4 posts)→

Original post →

More from coding & agent

coding & agent channel →