Apps & Tools
ModLens
A vision bridge for text-only coding agents that enables image pasting and OCR.
byliustack
Open source
ModLens media is blocked
Allow external media to connect to the provider and play this content.
Description
ModLens is a vision plugin designed for DeepSeek Harness and other text-only coding agents. It allows users to paste images directly into the chat, which the tool converts into structured JSON evidence—including OCR, layout regions, and entity lists—giving text-only models the ability to 'see' and ground their answers in visual data.
Descriptions, tags, and model credits may be AI-generated or inferred from public sources and can be incomplete or wrong. Learn more.