STARDUSTLC666/dsh-voice
DeepSeek Harness 语音双件套插件:voice_tts(edge-tts 协议免费微软神经语音合成,Sec-MS-GEC 本地 DRM 生成)/ voice_stt(OpenAI 兼容 ASR)/ voice_list。· 为 DeepSeek Harness 智能体提供 TTS + STT 功能。
Project Overview项目介绍
This is a natively built plugin exclusively for DeepSeek Harness that adds full speech interaction capabilities to DSH agents. It ships with three core tools: voice_tts for text-to-speech, voice_stt for speech-to-text, and voice_list to query available TTS voices. You can install it directly via the official DSH CLI with the command dsh plugin --profile web add dsh-voice, and it follows the DSH native bundle manifest specification for easy loading. It has been validated to work correctly on the official DSH v0.1.5-rc.1 release with Node.js 24.16.0, and supports over 22 common TTS voices across multiple languages.
This plugin is designed for DSH users who want to add voice interaction to their agent workflows. It fits common use cases like generating voice broadcasts for daily news or alerts, and transcribing pre-recorded meeting audio or voice notes into editable text. A typical workflow starts with querying the full voice list to pick your preferred speaking style, then calling the TTS tool to generate an MP3 file from any input text. The TTS tool works out of the box with no extra configuration required for most users, and for transcription, you simply pass the path to your local audio file to the STT tool to get results.
This plugin requires a Node.js version matching DSH’s requirements: 22.19 or newer for the 22.x branch, or any 24.x or newer release. The TTS functionality is completely free to use, with authentication tokens generated locally on your device, so no API key is required upfront. You can store your ASR API key as an environment variable for better security instead of putting it directly in the plaintext config file. The project is released under the permissive MIT open source license, with hard limits of 5000 characters per TTS request and 25MB per audio file for STT requests.
这是一款专为 DeepSeek Harness 开发的原生语音功能插件,为 DSH 代理新增语音交互能力,一共提供三个核心工具:依托微软 Edge 免费无限量神经语音服务的文字转语音工具、兼容 OpenAI 格式 Whisper 接口的语音转文字工具,以及可供查询的音色列表工具。它遵循 DSH 原生插件打包规范,可直接通过 DSH 官方 CLI 命令完成安装,适配官方 v0.1.5-rc.1 版本的加载规则。
这款插件适合需要让 DSH 代理实现语音交互的用户,可用于合成每日早报、系统通知的语音播报,也可对会议录音、用户语音留言做自动文字转写。典型工作流中,用户可先调用 voice_list 工具查看所有可用音色,选择偏好后调用 voice_tts 合成任意文本的 MP3 语音文件,也可上传本地音频文件调用 voice_stt 完成高精度转写。默认配置即可使用 TTS 功能,STT 仅需配置对应 ASR 接口的密钥即可启用。
该插件要求 Node.js 版本不低于 22.19 的 22.x 分支,或 24 及以上版本,和对应版本 DSH 的环境要求一致。TTS 完全免费无成本,令牌在本地生成,不需要额外申请密钥;STT 的成本由用户选择的 ASR 服务商收取,支持 Groq、OpenAI 以及自定义端点。项目采用 MIT 开源许可证,文本限制单次不超过 5000 字符,音频单次不超过 25MB。
请帮我安装这个 DSH 插件。安装前先完成【兼容性检查 + 安全性检查】,检查通过再动手。
插件:dsh-voice(STARDUSTLC666/dsh-voice)
仓库:https://github.com/STARDUSTLC666/dsh-voice
本站详情页:https://www.yhbd.top/plugins/stardustlc666-dsh-voice/
本站登记:类型 plugin · 归类 原生 DSH 插件 · 许可证 MIT · ⭐ 4 · 最近提交 2026-10-03 · 主语言 JavaScript
按下面顺序执行,每步先把结论告诉我,再进入下一步:
【1 兼容性检查】
① 我这边:DSH 版本、Node 版本、操作系统、当前 profile(web / desktop)。
② 读它的 README、package.json、插件 manifest,列出它要求的 DSH 版本 / Node 版本 / 操作系统 / 外部依赖 / 需要另外先装的运行时。
③ 逐条比对,结论只写「满足 / 不满足 / 未知」三种;不满足的给出可行替代方案。
④ 检查是否和我已装的插件冲突:命令名重复、skill / tool 重名、端口占用、重复注册的 MCP server。
【2 安全性检查】
① 仓库可信度:和上面「本站登记」是否一致;star / fork 数、创建时间、最近提交,是否归档或长期停更。
② 安装脚本:逐行看 package.json 的 preinstall / install / postinstall,以及 install.sh、setup.ps1 之类脚本。出现 curl|bash、下载后直接执行、混淆代码、访问与插件功能无关的域名,立刻停下来告诉我,不要继续装。
③ 依赖:列出新增依赖,标出无人维护、或与知名包拼写近似的可疑包(typosquatting)。
④ 权限与副作用:它会读写哪些目录、访问哪些域名、需要哪些 DSH 权限(filesystem / network / shell / clipboard 等),以及怎么卸载和回滚。
⑤ 如果它要求 sudo / 管理员权限,或权限明显超出功能所需,先停下来问我。
【3 安装】
上面两步没有「不满足」和「高危项」时才执行;用官方推荐方式安装,不要自行提权。
【4 汇报】
用表格输出:检查项 / 结论 / 依据 / 是否需要我决策。拿不准的一律写「未知」并说明要我怎么确认——不要猜,也不要替我决定。
Send this message to DSH in your current session: it verifies compatibility and security first (answering met / not met / unknown item by item) and only installs once everything checks out — it will stop and ask you if it finds a high-risk item. The box scrolls; the copy is the full prompt. CLI install commands may not be accurate across systems, so DSH is the safer route.把上面这条消息直接发给当前会话里的 DSH:它会先核对兼容性与安全性(逐条给「满足 / 不满足 / 未知」),确认没问题再安装,有高危项会停下来问你。框内可滚动,复制到的是完整提示词;安装命令不一定准确,发给 DSH 更稳。
- Only 4 stars - very few users, little community feedback星标只有 4,几乎没人在用,遇到问题缺少社区反馈
DSH walks through these 9 checksDSH 会逐条核对这 9 项
Compatibility兼容性
- DSH, Node, OS and profile requirementsDSH 版本 / Node 版本 / 操作系统 / profile 是否满足要求
- External dependencies and runtimes (Electron / Python / Docker, ...)外部依赖与运行时(Electron / Python / Docker 等)是否齐备
- Conflicts with installed plugins: command names, skill / tool names, ports, duplicate MCP registration与已装插件是否冲突:命令名、skill / tool 重名、端口占用、重复 MCP 注册
Security安全性
- Repo matches the facts registered here; archived or abandoned?仓库是否与页面登记一致,是否归档或长期停更
- Safety of preinstall / install / postinstall and install.sh / setup.ps1preinstall / install / postinstall 与 install.sh、setup.ps1 是否安全
- curl|bash, download-then-execute, obfuscation, unrelated domains → stop immediatelycurl|bash、下载即执行、混淆代码、无关域名 → 立刻停止
- Typosquatting or unmaintained packages among the new dependencies新增依赖里有没有 typosquatting 或无人维护的包
- Requested permissions vs. what the feature actually needs申请了哪些权限、是否超出功能所需(filesystem / network / shell / clipboard)
- Any sudo / admin requirement, plus uninstall and rollback是否要求 sudo / 管理员权限,以及卸载与回滚方式
Anything uncertain must be marked unknown with a note on how to confirm it. This site's signal screen is a static snapshot, not a security audit.拿不准的必须标「未知」并说明要我怎么确认。本站的信号筛查是静态快照,不能替代安全审计。
Or use CLI install (for developers)或使用命令行安装(适合开发者)
CLI Install命令行安装
dsh plugin --profile desktop add dsh-voice
把 STARDUSTLC666/dsh-voice 加入你的 DSH 配置(web profile)即可启用。
READMEREADME
dsh-voice

把文字生成语音,或通过兼容接口把音频转为文字。
功能
- 使用 Edge 在线朗读服务生成语音。
- 列出音色并批量试听。
- 通过 OpenAI 兼容 ASR 接口转写音频。
安装
桌面版可在「插件」面板按包名 dsh-voice 安装。已配置 dsh 命令时也可使用:
dsh plugin --profile desktop add dsh-voice
网页版把命令中的 desktop 改为 web。安装后重启 DSH。
开始使用
可说:“把这段文字生成中文 MP3,并给我两个音色试听。”配置 ASR 后,也可要求转写音频文件。
依赖与配置
语音合成需要网络连接;转写需配置对应 ASR 服务与密钥。代理可以单独设置。
详细配置、工具参数与排错见使用说明。从源码独立开发时,Node 要求以 package.json 为准。
Totoro-qaq/dsh-plugin-bridge
william-jin-cmu/dsh-vision
Flyvhidbwo/dsh-vision-proxy
WNJXYK/dsh-codex-oauth
JuneLearn/dsh-image2-draw
mjylfz/dsh-skill-mover
starefinger/dsh-llm-qwen-local
moduqishi/GrassVison