← Back to list

dsh-auto-vision

DeepSeek Harness Other Low risk

Auto-discovery vision bridge for text-only DeepSeek Harness agents: automatically finds an image-capable model from your configured providers and returns picture descriptions as plain text via a vision tool.

为纯文本 DeepSeek Harness 主模型搭视觉桥:自动从已配置模型中挑选多模态模型,通过 vision 工具把图片描述以纯文本返回。

How to install

DeepSeek Harness dsh plugin add dsh-auto-vision

About

dsh-auto-vision **给 DeepSeek Harness 里的纯文本主模型装上眼睛:自动发现你已配置的多模态模型,一条命令装上 vision 工具,图片识别结果以纯文本返回。** 快速开始 本插件**已发布到 npm**,两种安装方式任选: **方式一:npm 安装(推荐)** sh dsh plugin --profile <你的profile名> add dsh-auto-vision **方式二:GitHub 源码安装**(纯 JS、零构建步骤,无需构建授权) sh dsh plugin --profile <你的profile名> add github:NormanFxxkingRockwell/dsh-auto-vision 装好后,直接在主对话里说: 读这张图 C:\path\to\image.jpg 描述一下 主模型会自动调用 vision 工具,把识别结果以文本形式返回给你。 要求:你的 dsh 里已经配置了至少一个**声明了图片输入**的多模态模型(如何声明见下文「配置」)。没有的话,插件会在启动时报错并告诉你怎么办。 它解决什么问题 dsh 内置的 read_image 会把图片块直接塞进当前模型的上下文,所以只有当**当前主模型本身支持图片**时才能用。像 deepseek v4 flash 这样的纯文本模型,调用 read_image 会被直…

Recommendation signals

48 Tool quality · Based on stars, downloads, maintenance, security and docs
– User interest · Adjusted by in-site views, install copies and download clicks
48 Overall
0views
0unique visitors
0install copies
0download clicks
0outbound clicks

Meta

License
MIT
Language
JavaScript
GitHub stars
1
mo. downloads
840
Last push
2026-09-15
Created
2026-08-17

Links

Basic safety check

Findings
None
Sources
curated:awesome-dsh-plugin.com, curated:awesome-dsh-plugin/awesome-dsh-plugin
Topics
cordic, deepseek-harness, dsh, dsh-plugin, image-to-text, multimodal, vision

Related plugins

DeepSeek HarnessClaude CodeCodex
Score92

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

☆ 4.0K ↓ 106.1K Other ↗
DeepSeek Harness Featured
Score72

dsh-web

zhu1090093659/dsh-web-ui/tree/main/packages/dsh-tool-describe-image

A `describe_image` vision tool for text-only models: images (local path, URL, attachment) go to a configurable OpenAI-compatible vision endpoint and only the returned text enters the session.

☆ 8.0K ↓ – Other ↗
DeepSeek Harness
Score70

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

☆ 1.1K ↓ 66.0K Other ↗
DeepSeek Harness
Score67

dsh-vision-toolkit

Anionex/dsh-vision-toolkit

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

☆ 883 ↓ 42.3K Other ↗