Apps & Tools
ModLens
A vision bridge for text-only coding agents that enables image pasting and OCR.
byliustack
Open source
Description
ModLens is a vision plugin designed for DeepSeek Harness and other text-only coding agents. It allows users to paste images directly into the chat, which the tool converts into structured JSON evidence—including OCR, layout regions, and entity lists—giving text-only models the ability to 'see' and ground their answers in visual data.
Descriptions, tags, and model credits may be AI-generated or inferred from public sources and can be incomplete or wrong. Learn more.