❯ DeepSeek Harness Ships rc.8, Adding Image Input and Installable Sub-Agents
RELEASEDeepSeek Harness released v0.1.0-rc.8 on August 19, with 14 changes total, centered on rounding out multimodal input: the model adaptation layer now supports native image requests, core commands like /goal and /plan accept mixed image-and-text input for the first time, and the @ menu adds references to local files and past sessions.
DETAILSThe sub-agent system was refactored into optional installable packages, so both Claude Code and Codex can be installed and invoked. More noteworthy is the “tool vision” layer built for text-only models — it converts images into structured information via OCR, color statistics, and pixel scanning, letting models without vision capabilities reason over screenshots. The release also fixes long-standing issues with image requests, streaming generation, and custom gateways, and improves the Windows terminal experience. Per the official changelog, plugin support was strengthened too, adding concurrent web_search queries.
USAGEScreenshots, design mockups, and error screens can now be dropped straight into the command line, saving developers the step of describing interfaces in words — a step that was both time-consuming and prone to information loss. This tool-layer vision approach also shows that multimodality doesn’t have to wait for the model itself to upgrade: with OCR and pixel scanning paving the way, text-only models can take on image tasks too.
▪ SIGNALConverting images to text and feeding the result to text-only models paves the way with engineering before model capabilities catch up.