#663 · Primary category: AI Agents & Automation

dsh-vision-router

deepseek-harness dsh dsh-plugin multimodal vision

Eyes for text-only DeepSeek Harness agents: built-in free vision chain (no key) + pixel-level vision tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots). One-command install, no Python, image turns work like ordinary tool-calling turns.

Project last updated:08/29/26

GitHub Stars

1.0K

Forks

44

Contributors

6

License

MIT

Why we included this project

DeepSeek Harness agents only handle text, so anything visual, like a screenshot or a scanned page, is invisible to them. dsh-vision-router gives those agents a working pair of eyes: a free vision chain that runs without an API key, plus pixel-level tools for OCR, image Q&A, cropping, color analysis, SVG tracing, cutouts, and screenshots. It installs with one command and needs no Python, so you can add image understanding to an existing agent setup without spinning up extra services. Image turns work the same way ordinary tool calls do, which keeps the results easy to verify and repeat. If your DeepSeek Harness agents keep stalling on visual input, this is a quick way to let them see.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category