STARDUSTLC666/dsh-voice 预览 preview

STARDUSTLC666/dsh-voice

DeepSeek Harness 语音双件套插件:voice_tts(edge-tts 协议免费微软神经语音合成,Sec-MS-GEC 本地 DRM 生成)/ voice_stt(OpenAI 兼容 ASR)/ voice_list。· 为 DeepSeek Harness 智能体提供 TTS + STT 功能。

Project Overview项目介绍

This is a natively built plugin exclusively for DeepSeek Harness that adds full speech interaction capabilities to DSH agents. It ships with three core tools: voice_tts for text-to-speech, voice_stt for speech-to-text, and voice_list to query available TTS voices. You can install it directly via the official DSH CLI with the command dsh plugin --profile web add dsh-voice, and it follows the DSH native bundle manifest specification for easy loading. It has been validated to work correctly on the official DSH v0.1.5-rc.1 release with Node.js 24.16.0, and supports over 22 common TTS voices across multiple languages.

This plugin is designed for DSH users who want to add voice interaction to their agent workflows. It fits common use cases like generating voice broadcasts for daily news or alerts, and transcribing pre-recorded meeting audio or voice notes into editable text. A typical workflow starts with querying the full voice list to pick your preferred speaking style, then calling the TTS tool to generate an MP3 file from any input text. The TTS tool works out of the box with no extra configuration required for most users, and for transcription, you simply pass the path to your local audio file to the STT tool to get results.

This plugin requires a Node.js version matching DSH’s requirements: 22.19 or newer for the 22.x branch, or any 24.x or newer release. The TTS functionality is completely free to use, with authentication tokens generated locally on your device, so no API key is required upfront. You can store your ASR API key as an environment variable for better security instead of putting it directly in the plaintext config file. The project is released under the permissive MIT open source license, with hard limits of 5000 characters per TTS request and 25MB per audio file for STT requests.

这是一款专为 DeepSeek Harness 开发的原生语音功能插件,为 DSH 代理新增语音交互能力,一共提供三个核心工具:依托微软 Edge 免费无限量神经语音服务的文字转语音工具、兼容 OpenAI 格式 Whisper 接口的语音转文字工具,以及可供查询的音色列表工具。它遵循 DSH 原生插件打包规范,可直接通过 DSH 官方 CLI 命令完成安装,适配官方 v0.1.5-rc.1 版本的加载规则。

这款插件适合需要让 DSH 代理实现语音交互的用户,可用于合成每日早报、系统通知的语音播报,也可对会议录音、用户语音留言做自动文字转写。典型工作流中,用户可先调用 voice_list 工具查看所有可用音色,选择偏好后调用 voice_tts 合成任意文本的 MP3 语音文件,也可上传本地音频文件调用 voice_stt 完成高精度转写。默认配置即可使用 TTS 功能,STT 仅需配置对应 ASR 接口的密钥即可启用。

该插件要求 Node.js 版本不低于 22.19 的 22.x 分支,或 24 及以上版本,和对应版本 DSH 的环境要求一致。TTS 完全免费无成本,令牌在本地生成,不需要额外申请密钥;STT 的成本由用户选择的 ASR 服务商收取,支持 Groq、OpenAI 以及自定义端点。项目采用 MIT 开源许可证,文本限制单次不超过 5000 字符,音频单次不超过 25MB。

Pre-install check安装前体检Compatibility · Security兼容性 · 安全性 1 warning1 项注意
  • Only 4 stars - very few users, little community feedback星标只有 4,几乎没人在用,遇到问题缺少社区反馈
DSH walks through these 9 checksDSH 会逐条核对这 9 项

Compatibility兼容性

  • DSH, Node, OS and profile requirementsDSH 版本 / Node 版本 / 操作系统 / profile 是否满足要求
  • External dependencies and runtimes (Electron / Python / Docker, ...)外部依赖与运行时(Electron / Python / Docker 等)是否齐备
  • Conflicts with installed plugins: command names, skill / tool names, ports, duplicate MCP registration与已装插件是否冲突:命令名、skill / tool 重名、端口占用、重复 MCP 注册

Security安全性

  • Repo matches the facts registered here; archived or abandoned?仓库是否与页面登记一致,是否归档或长期停更
  • Safety of preinstall / install / postinstall and install.sh / setup.ps1preinstall / install / postinstall 与 install.sh、setup.ps1 是否安全
  • curl|bash, download-then-execute, obfuscation, unrelated domains → stop immediatelycurl|bash、下载即执行、混淆代码、无关域名 → 立刻停止
  • Typosquatting or unmaintained packages among the new dependencies新增依赖里有没有 typosquatting 或无人维护的包
  • Requested permissions vs. what the feature actually needs申请了哪些权限、是否超出功能所需(filesystem / network / shell / clipboard)
  • Any sudo / admin requirement, plus uninstall and rollback是否要求 sudo / 管理员权限,以及卸载与回滚方式

Anything uncertain must be marked unknown with a note on how to confirm it. This site's signal screen is a static snapshot, not a security audit.拿不准的必须标「未知」并说明要我怎么确认。本站的信号筛查是静态快照,不能替代安全审计。

Or use CLI install (for developers)或使用命令行安装(适合开发者)

CLI Install命令行安装

dsh plugin --profile desktop add dsh-voice

把 STARDUSTLC666/dsh-voice 加入你的 DSH 配置(web profile)即可启用。

READMEREADME

dsh-voice

English

dsh-voice 鲸鱼娘插件封面

把文字生成语音,或通过兼容接口把音频转为文字。

npm downloads

功能

  • 使用 Edge 在线朗读服务生成语音。
  • 列出音色并批量试听。
  • 通过 OpenAI 兼容 ASR 接口转写音频。

安装

桌面版可在「插件」面板按包名 dsh-voice 安装。已配置 dsh 命令时也可使用:

dsh plugin --profile desktop add dsh-voice

网页版把命令中的 desktop 改为 web。安装后重启 DSH。

开始使用

可说:“把这段文字生成中文 MP3,并给我两个音色试听。”配置 ASR 后,也可要求转写音频文件。

依赖与配置

语音合成需要网络连接;转写需配置对应 ASR 服务与密钥。代理可以单独设置。

详细配置、工具参数与排错见使用说明。从源码独立开发时,Node 要求以 package.json 为准。

文档

License

MIT

← 上一个 Prev dsh-stream-rules 下一个 Next dsh-mcp-admin →