What it is — ModLens bridges text-only models with vision, allowing pasted images to obtain structured JSON evidence including OCR, layout, and semantics.
Who it is for — When using DeepSeek Harness with text-only models like DeepSeek-V4-Flash to analyze pasted images, it is suitable.
For users of native multimodal models, this plugin is not necessary because native models retain their built-in image pasting capability.
Watch out — Sandbox testing passed, with successful installation in a fresh profile and registration with Harness. No obvious issues found.
The verdict — I would install it because it outputs structured visual evidence without saving the image to a file path first.