← Back to list

dsh-guide-dog

DeepSeek Harness Other Low risk

MiniMax-powered multimodal plugin: real-time voice call mode (streaming conversation, floating dock UI), voice mode and mic voice input, plus image/video/music/speech generation and vision inspection tools.

基于 MiniMax 的多模态插件:实时语音通话模式(流式对话、悬浮胶囊 UI)、语音模式与麦克风语音输入,并提供图像/视频/音乐/语音生成与视觉检查工具。

How to install

DeepSeek Harness dsh plugin add github:AtropinolTT/dsh-guide-dog

About

Guide Dog for DSH, powered by MiniMax **English** | 简体中文 A dynamic Cordis plugin that gives DeepSeek Harness multimodal superpowers through the mmx CLI (MiniMax): **Eyes for DeepSeek** — MiniMax VLM (guide_dog_vision / guide_dog_inspect) describes images, so a model with no native vision input (e.g. DeepSeek) can still review frontend designs, figures, screenshots, and generated images. **Hands for generation** — images (image-01), video (MiniMax-H3 / Hailuo), speech (MiniMax TTS), music (music-3.0), text (MiniMax-M3), and web search. **Web UI preview & playback** — every generated file is ser…

Recommendation signals

42 Tool quality · Based on stars, downloads, maintenance, security and docs
– User interest · Adjusted by in-site views, install copies and download clicks
42 Overall
0views
0unique visitors
0install copies
0download clicks
0outbound clicks

Meta

License
MIT
Language
JavaScript
GitHub stars
5
mo. downloads
–
Last push
2026-08-17
Created
2026-08-14

Links

GitHub ↗ Report issue ↗

Basic safety check

Findings
curated 收录但无 npm 包/安装命令
Sources
curated:awesome-dsh-plugin.com, curated:awesome-dsh-plugin/awesome-dsh-plugin
Topics
dsh, dsh-plugin, minimax, multimodal, tts, voice-input

Related plugins

DeepSeek HarnessClaude CodeCodex
Score92

modlens

liustack/modlens

Vision bridge for text-only models: paste an image, get structured JSON evidence (OCR, layout, semantics).

☆ 4.0K ↓ 106.1K Other ↗
DeepSeek Harness Featured
Score72

dsh-web

zhu1090093659/dsh-web-ui/tree/main/packages/dsh-tool-describe-image

A `describe_image` vision tool for text-only models: images (local path, URL, attachment) go to a configurable OpenAI-compatible vision endpoint and only the returned text enters the session.

☆ 8.0K ↓ – Other ↗
DeepSeek Harness
Score70

dsh-vision-router

ysr666/dsh-vision-router

Free vision for text-only agents: built-in keyless vision chain plus pixel tools (Q&A, grounding, crop, pixel diff, colors, OCR, SVG trace, cutout, screenshots); paste an image to use it.

☆ 1.1K ↓ 66.0K Other ↗
DeepSeek Harness
Score67

dsh-vision-toolkit

Anionex/dsh-vision-toolkit

Vision for text-only models: paste an image and the model switches to a Vision Toolkit variant for image Q&A, multi-image comparison, long-screenshot OCR, screenshot-to-UI reproduction, element grounding, and pixel diff. No API key by default — images are processed by the author-hosted free service, 100 per machine per day; configurable to your own provider.

☆ 883 ↓ 42.3K Other ↗