liustack/modlens: The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). 全网最强 Deep
Why CEREBRO kept it
Vision plugin extends text-only agents with multimodal OCR/layout capability
The text below is an automated extraction of the article at https://github.com/liustack/modlens, stored verbatim in the public cerebro-vault repository. Copyright remains with the original publisher (github.com).
Give a text-only model sight, and just paste the image. 🥇 The most capable vision plugin for DeepSeek Harness (dsh) 🥇 简体中文 · Configuration · Troubleshooting · Security · 🔍 ModSearch (the best free web search plugin for DSH) The flagship DeepSeek and GLM chat models are text-only and cannot read images. ModLens is a plug-in vision engine that gives a text-only model sight. ModLens reads images pasted straight into the chat, no saving to a file and passing a path first. Issues are welcome any time: open one. Follow the liustack WeChat official account, and come find me on X: @liustack. What you
Backlinks
Appeared in 1 briefing