dsh-vision
Vision for text-only DeepSeek via Doubao Web by default (zero-cost, no API key — drives your logged-in Chrome through a Windows CDP bridge), with Antigravity IDE quota (flash/pro) or Gemini fallback; auto detail escalation, vision evidence memory with compaction rehydration, content-hash cache, and a bilingual client panel.
纯文本 DeepSeek 的识图插件:默认走豆包 Web(零成本、免 API key,经 Windows CDP 桥接驱动已登录 Chrome),另有反重力额度(flash/pro)与 Gemini 降级通道;档位自动升级、视觉证据记忆与会话压缩恢复、内容哈希缓存、双语客户端面板。
How to install
dsh plugin add dsh-vision-web About
👁️ dsh-vision **给 DeepSeek Harness 的纯文本 DeepSeek 一双眼睛——默认走豆包 Web,零成本、免 API key。** English | 简体中文 **一句话**:不用 API key、不用付费——浏览器里登录一次豆包,DeepSeek 就能在每轮对话里看图。粘贴、识图、回答。 ✨ 为什么值得用 | 痛点 | dsh-vision 的解法 | |---|---| | 视觉 API 花钱还要 key | **豆包 Web 默认通道**——零成本、免 API key,浏览器登录即可 | | DeepSeek 纯文本,粘贴图片被拒 | 包装适配器声明图片输入,图片自动转文本占位 | | 别的插件锁死一家厂商 | 豆包 Web(默认)+ 反重力额度(flash/pro)+ Gemini API + Cockpit 反代——自动降级链 | | 模型看到图但"忘了" | **视觉证据记忆**:结果持久化在会话,跨轮复用,压缩后恢复 | | 同一张图反复花钱识别 | **内容哈希缓存**:同图同问每进程最多识别一次 | | 复杂画面识别不准 | **档位自动升级**:先标准检查,复杂画面自动深度检查 | | WSL / 被墙环境 | API 通道有 winCurl 降级;豆包通道走你的 Windows 浏览器 | 🎯 真实效果(2026-08 实…
Recommendation signals
Meta
- License
- MIT
- Language
- TypeScript
- GitHub stars
- 7
- mo. downloads
- 760
- Last push
- 2026-08-16
- Created
- 2026-08-16
Basic safety check
- Findings
- None
- Sources
- curated:awesome-dsh-plugin.com, curated:awesome-dsh-plugin/awesome-dsh-plugin
- Topics
- antigravity, doubao, dsh, dsh-plugin, gemini, image-recognition, multimodal, ocr, vision, vlm
Related plugins
modlens
liustack/modlens
Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).
dsh-web
zhu1090093659/dsh-web-ui/tree/main/packages/dsh-tool-describe-image
A `describe_image` vision tool for text-only models: images (local path, URL, attachment) go to a configurable OpenAI-compatible vision endpoint and only the returned text enters the session.
dsh-vision-router
ysr666/dsh-vision-router
Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.
dsh-vision-toolkit
Anionex/dsh-vision-toolkit
Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.