← Back to list

dsh-vision-toolkit

DeepSeek Harness Other Low risk

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

让纯文本模型处理视觉任务:粘贴图片后自动切换到 Vision Toolkit 变体,支持图片问答、多图比较、长截图 OCR、截图还原前端 UI、元素定位与像素对比。默认无需 API Key——图片经作者自建的免费服务处理,每台机器每天 100 张;也可改为指向自己的服务商。

How to install

DeepSeek Harness dsh plugin add @anionex/dsh-vision-toolkit

About

DSH Vision Toolkit **A more powerful vision toolkit—give text-only models in DeepSeek Harness eyes: image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks in one toolkit and Skill.** 🚀 Paste an image and ask directly | Install with one command | Built-in free vision | Broad use cases 🌐 **English** | 中文 If you use DeepSeek or another text-only model in DeepSeek Harness (DSH), you may have run into the same problems: the model cannot see a screenshot, generic descriptions miss the point, buttons have no usable coordinates, and a rebuilt page can look “close enough” without a way …

Recommendation signals

67 Tool quality · Based on stars, downloads, maintenance, security and docs
– User interest · Adjusted by in-site views, install copies and download clicks
67 Overall
0views
0unique visitors
0install copies
0download clicks
0outbound clicks

Meta

License
MIT
Language
TypeScript
GitHub stars
883
mo. downloads
42.3K
Last push
2026-09-16
Created
2026-08-13

Links

Basic safety check

Findings
None
Sources
curated:awesome-dsh-plugin.com, curated:awesome-dsh-plugin/awesome-dsh-plugin, curated:0xsline/awesome-deepseek-harness
Topics
agent-skills, agent-vision-toolkit, computer-vision, deepseek, deepseek-harness, dsh, dsh-plugin, gui-automation, ocr, plugin, python, screenshot-testing, text-only-llm, typescript, ui-restoration, vision-language-model, vision-tools

Related plugins

DeepSeek HarnessClaude CodeCodex
Score92

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

☆ 4.0K ↓ 106.1K Other ↗
DeepSeek Harness Featured
Score72

dsh-web

zhu1090093659/dsh-web-ui/tree/main/packages/dsh-tool-describe-image

A `describe_image` vision tool for text-only models: images (local path, URL, attachment) go to a configurable OpenAI-compatible vision endpoint and only the returned text enters the session.

☆ 8.0K ↓ – Other ↗
DeepSeek Harness
Score70

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

☆ 1.1K ↓ 66.0K Other ↗
DeepSeek Harness
Score67

dsh-design-qa

sunxin-ai/dsh-design-qa

Design-fidelity QA for text-only models: a `deepseek_vision` tool borrows an eye from any OpenAI-compatible vision route, so the model can judge whether an implementation matches its mock — shipped with the benchmark behind that judgement (four fixtures, 23 injected defects, raw transcripts) and the questioning discipline it depends on.

☆ 44 ↓ 560 Other ↗