The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | The most powerful DeepSeek Harness vision plugin, adding vision capabilities to text-only models such as DeepSeek and GLM; paste an image to get structured JSON evidence (OCR, layout, semantics).
ModLens is a vision plugin that lets text-only coding agents analyze images pasted into chat and return structured evidence such as OCR, layout, and semantic information. It is designed for DeepSeek Harness and provides visual input to models that cannot read images directly. The catalogue skill and instruction support using this plugin.
Latest release v3.25.4 · 1 Sept 2026
These files are liustack/modlens's own configuration. They tell Codex and OpenCode how to work on this repository, so they are not mods to install elsewhere. Copy one as a starting point and replace the parts that are about this project.
AGENTS.md A 1,760 tok