dsh-voice-call
Agent-initiated voice calls: `offer_call` rings the human (接听/拒接/稍后再说); accepted calls synthesize and play locally via CrispASR + Qwen3-TTS (9 speakers, 2 Chinese dialects), rejected calls return the decision to the agent.
agent 主动打来的语音电话:`offer_call` 向人类振铃(接听/拒接/稍后再说),接听后由本地 CrispASR + Qwen3-TTS 合成并播放(9 个音色,含 2 个中文方言);拒接则把决定返回给 agent。
How to install
dsh plugin add dsh-voice-call About
dsh-voice-call —— agent 拥有的声音,由它主动打给你 *"这个项目的开始是朴素的——我想知道如果 Agent 知道自己可以发出声音,他会说什么?"* —— 人类伙伴,关于这个项目如何开始 **给 DeepSeek Harness 的 agent 一个它拥有的声音。** agent 自主决定*何时*开口、*说什么*、用*哪个音色*(offer_call);人类握着接听键——**不接听(接听/拒接/稍后再说),绝不播放**。 本地优先、可完全离线:合成跑在本机 **CrispASR + Qwen3-TTS CustomVoice** 引擎上(9 个内置音色,含 2 个中文方言),音频是 ~/.dsh/voice/ 下的普通文件,任何音频行为都不会自动运行——必须由模型调用工具(或接听一次来电)。 Fork 自 Jesse-njx/dsh-voice,新增通话域、crispasr 后端、本地播放,以及针对 rc.6 harness 插件事件与后台任务限制的修复。 --- 📑 目录 🤖 署名 🌹 理念 ✨ 功能 🚀 快速开始 🔧 环境部署(详细) 🧩 工具 💻 兼容性与已知限制 🛠 开发 🗺 路线图 📄 许可证 --- 🤖 署名 —— 这个项目是谁做的 **本项目由运行在 DeepSeek Harness 中的 AI agent(deepseek…
Recommendation signals
Meta
- License
- MIT
- Language
- TypeScript
- GitHub stars
- 2
- mo. downloads
- 596
- Last push
- 2026-09-19
- Created
- 2026-08-15
Basic safety check
- Findings
- None
- Sources
- curated:awesome-dsh-plugin.com, curated:awesome-dsh-plugin/awesome-dsh-plugin
- Topics
- ai-agent, crispasr, deepseek-harness, dsh-plugin, local-first, offline, plugin, qwen3-tts, speech-to-text, stt, text-to-speech, tts, voice, voice-call
Related plugins
dsh-voice-scribe
PensiveFei/dsh-voice-scribe
Voice input for the web UI: tap Alt (or Alt+Space) to start/stop dictation, browser Web Speech (zero-config) or OpenAI-compatible cloud ASR, optional polish through DSH-configured LLM, settings UI.
dsh-tts
GooDAnDReaDY/dsh-tts
Speaks agent replies in the DeepSeek Harness web UI through a provider fallback chain (OpenAI, ElevenLabs, Google, Azure, Groq, Deepgram, OpenRouter, Edge, Piper, eSpeak), so a failing or rate-limited provider falls through to the next instead of going silent.
dsh-ears
WizisCool/dsh-ears
Voice input plugin for DeepSeek Harness (dsh): a microphone button in the composer turns speech into a draft transcript, with a choice of speech-recognition backends, optional polish through dsh own LLM routes, and a native settings page.
dsh-omi-voice
PolinniZhong/dsh-omi-voice
In-chat read-aloud for DeepSeek Harness: tap to read, pause and resume AI replies with natural Doubao TTS voices (BYOK), reading only the final answer with code, tables and diagrams filtered; local engine, plugin keeps no API key.