← Back to list

dsh-vision-proxy

DeepSeek Harness Tools & Capabilities High risk

DeepSeek brain + automatic image transcription: attach images in the GUI and each one is transcribed to text via any OpenAI-compatible VLM before reaching the text-only DeepSeek — a keyed fast path (default qwen3.7-flash; DashScope/Zhipu/OpenRouter or any OpenAI-compatible endpoint) with your own key, or local Ollama auto-detected with zero config.

DeepSeek 大脑 + 自动识图:GUI 附加的每张图片自动经 OpenAI 兼容 VLM 转译成文字,再交给纯文本的 DeepSeek 作答——有 key 自动走快速通道(默认 qwen3.7-flash,支持百炼/智谱/OpenRouter 等任意 OpenAI 兼容端点),无 key 自动探测本地 Ollama(零配置,图片不出本机)。

How to install

DeepSeek Harness dsh plugin add dsh-vision-proxy

Recommendation signals

54 Tool quality · Based on stars, downloads, maintenance, security and docs
User interest · Adjusted by in-site views, install copies and download clicks
54 Overall
0views
0install copies
0download clicks
0outbound clicks

Meta

License
MIT
Language
JavaScript
GitHub stars
8
mo. downloads
Last push
2026-08-14
Created
2026-08-13

Links

Basic safety check

Findings
package.json 含 postinstall 脚本
Sources
curated:awesome-dsh-plugin.com, curated:awesome-dsh-plugin/awesome-dsh-plugin, curated:0xsline/awesome-deepseek-harness
Topics
dashscope, deepseek-harness, dsh-plugin, image-understanding, multimodal, ocr, qwen, vision, vlm

Related plugins

DeepSeek HarnessClaude CodeCodex Featured
Score72

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

☆ 1.7K ↓ – Tools & Capabilities
DeepSeek Harness Featured
Score67

dsh-vision-toolkit

Anionex/dsh-vision-toolkit

Vision tasks for text-only models: intent-aware image Q&A, long-screenshot OCR, UI reproduction, grounding, and pixel diff.

☆ 403 ↓ – Tools & Capabilities
DeepSeek Harness
Score64

dsh-browser

Lum1104/dsh-browser

Chrome sidebar extension that lets DSH operate your browser directly, no vision capabilities required.

☆ 133 ↓ – Tools & Capabilities
DeepSeek Harness
Score63

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

☆ 108 ↓ – Tools & Capabilities