FuzzySoul/dsh-chatvoice
ChatVoice — DeepSeek Harness (dsh) 的免费语音输入与 AI 回复朗读插件。零配置/零成本/免 API key 语音插件。
Project Overview项目介绍
This is a native plugin built exclusively for the DeepSeek Harness web client, adding full voice functionality including speech-to-text for prompts and text-to-speech for AI responses. It relies entirely on the browser’s native Web Speech API, so no third-party API keys are required, there is no cost to use it, and you don’t need to do any extra configuration after installation. To install the plugin, run the command dsh plugin --profile web add dsh-chatvoice in your terminal, then restart the DSH web service and navigate to the local address 127.0.0.1:3080 to get started.
The plugin adds a microphone button next to the input box for voice input. Once you click the button and grant mic permission to your browser, confirmed speech results are added to the input box in real time, while intermediate results show in a temporary bubble at the top of the interface. You can manually edit or delete text in the input box while recognition is still running, and the plugin only appends new speech content instead of overwriting your edits.
You can customize the recognition language, auto-read toggle, voice, and speech rate in the DSH settings menu under the ChatVoice section, and changes take effect immediately without requiring a restart of DSH. For users in mainland China, the Edge browser is recommended, as it offers more stable speech recognition via Azure and access to the free natural-sounding Xiaoxiao Online Mandarin voice. Firefox and Safari do not support the SpeechRecognition API required for voice input, so only the text-to-speech feature works on these browsers.
这是专为 DeepSeek Harness (DSH) 开发的原生插件,为 DSH Web 版本添加了语音输入和 AI 回复文本朗读的完整语音功能闭环。它全程调用浏览器原生的 Web Speech API,无需申请第三方 API 密钥,零配置开箱即用,也没有任何额外成本。安装时只需要在命令行执行 dsh plugin --profile web add dsh-chatvoice,重启 DSH Web 服务后访问本地地址即可使用。
核心功能包含语音输入和回复朗读两大块,语音输入会在输入框旁添加麦克风按钮,点击开始识别后,确认内容逐句实时进入输入框,中间结果显示在顶部气泡,识别过程中随时可以手动编辑修改内容,语音只会追加不会回写已有内容。每条 AI 回复旁都有朗读按钮,点击即可一键朗读,随时点击打断,还支持开启新回复自动朗读功能。
插件支持自定义识别语言、朗读音色、语速,设置保存后立即生效无需重启。国内网络环境下更推荐使用 Edge 浏览器,语音识别走 Azure 服务比 Chrome 更稳定,还能获得自然流畅的免费中文在线音色。受浏览器 API 限制,Firefox 和 Safari 不支持语音识别,只能使用朗读功能,访问必须用 127.0.0.1,LAN IP 访问会被禁用麦克风。
请帮我安装这个 DSH 插件。安装前先完成【兼容性检查 + 安全性检查】,检查通过再动手。
插件:dsh-chatvoice(FuzzySoul/dsh-chatvoice)
仓库:https://github.com/FuzzySoul/dsh-chatvoice
本站详情页:https://www.yhbd.top/plugins/fuzzysoul-dsh-chatvoice/
本站登记:类型 plugin · 归类 原生 DSH 插件 · 许可证 MIT · ⭐ 4 · 最近提交 2026-08-16 · 主语言 JavaScript
按下面顺序执行,每步先把结论告诉我,再进入下一步:
【1 兼容性检查】
① 我这边:DSH 版本、Node 版本、操作系统、当前 profile(web / desktop)。
② 读它的 README、package.json、插件 manifest,列出它要求的 DSH 版本 / Node 版本 / 操作系统 / 外部依赖 / 需要另外先装的运行时。
③ 逐条比对,结论只写「满足 / 不满足 / 未知」三种;不满足的给出可行替代方案。
④ 检查是否和我已装的插件冲突:命令名重复、skill / tool 重名、端口占用、重复注册的 MCP server。
【2 安全性检查】
① 仓库可信度:和上面「本站登记」是否一致;star / fork 数、创建时间、最近提交,是否归档或长期停更。
② 安装脚本:逐行看 package.json 的 preinstall / install / postinstall,以及 install.sh、setup.ps1 之类脚本。出现 curl|bash、下载后直接执行、混淆代码、访问与插件功能无关的域名,立刻停下来告诉我,不要继续装。
③ 依赖:列出新增依赖,标出无人维护、或与知名包拼写近似的可疑包(typosquatting)。
④ 权限与副作用:它会读写哪些目录、访问哪些域名、需要哪些 DSH 权限(filesystem / network / shell / clipboard 等),以及怎么卸载和回滚。
⑤ 如果它要求 sudo / 管理员权限,或权限明显超出功能所需,先停下来问我。
【3 安装】
上面两步没有「不满足」和「高危项」时才执行;用官方推荐方式安装,不要自行提权。
【4 汇报】
用表格输出:检查项 / 结论 / 依据 / 是否需要我决策。拿不准的一律写「未知」并说明要我怎么确认——不要猜,也不要替我决定。
Send this message to DSH in your current session: it verifies compatibility and security first (answering met / not met / unknown item by item) and only installs once everything checks out — it will stop and ask you if it finds a high-risk item. The box scrolls; the copy is the full prompt. CLI install commands may not be accurate across systems, so DSH is the safer route.把上面这条消息直接发给当前会话里的 DSH:它会先核对兼容性与安全性(逐条给「满足 / 不满足 / 未知」),确认没问题再安装,有高危项会停下来问你。框内可滚动,复制到的是完整提示词;安装命令不一定准确,发给 DSH 更稳。
- Only 4 stars - very few users, little community feedback星标只有 4,几乎没人在用,遇到问题缺少社区反馈
DSH walks through these 9 checksDSH 会逐条核对这 9 项
Compatibility兼容性
- DSH, Node, OS and profile requirementsDSH 版本 / Node 版本 / 操作系统 / profile 是否满足要求
- External dependencies and runtimes (Electron / Python / Docker, ...)外部依赖与运行时(Electron / Python / Docker 等)是否齐备
- Conflicts with installed plugins: command names, skill / tool names, ports, duplicate MCP registration与已装插件是否冲突:命令名、skill / tool 重名、端口占用、重复 MCP 注册
Security安全性
- Repo matches the facts registered here; archived or abandoned?仓库是否与页面登记一致,是否归档或长期停更
- Safety of preinstall / install / postinstall and install.sh / setup.ps1preinstall / install / postinstall 与 install.sh、setup.ps1 是否安全
- curl|bash, download-then-execute, obfuscation, unrelated domains → stop immediatelycurl|bash、下载即执行、混淆代码、无关域名 → 立刻停止
- Typosquatting or unmaintained packages among the new dependencies新增依赖里有没有 typosquatting 或无人维护的包
- Requested permissions vs. what the feature actually needs申请了哪些权限、是否超出功能所需(filesystem / network / shell / clipboard)
- Any sudo / admin requirement, plus uninstall and rollback是否要求 sudo / 管理员权限,以及卸载与回滚方式
Anything uncertain must be marked unknown with a note on how to confirm it. This site's signal screen is a static snapshot, not a security audit.拿不准的必须标「未知」并说明要我怎么确认。本站的信号筛查是静态快照,不能替代安全审计。
Or use CLI install (for developers)或使用命令行安装(适合开发者)
CLI Install命令行安装
dsh plugin --profile web add dsh-chatvoice
把 FuzzySoul/dsh-chatvoice 加入你的 DSH 配置(web profile)即可启用。
READMEREADME
ChatVoice 🎤🔊 — dsh-chatvoice
English | 中文
给 DeepSeek Harness (dsh) 装上「免费、免 API key、开箱即用」的语音输入 + AI 回复朗读闭环。 全程浏览器原生 Web Speech API —— 零配置、零成本、无任何后端与注册。

🎤 语音输入:确认句逐句实时入框,中间结果进气泡

🔊 回复朗读:点小喇叭一键朗读,可随时打断

✏️ 边听边改:聆听中打字修改/全删,语音只追加不回写
ChatVoice = Chat + Voice:一个插件解决「嘴」和「耳朵」——写代码时手不离键盘,用嘴问 AI;懒得看长回复,让 AI 读给你听(听力型学习 / 无障碍 / 摸鱼躺用场景全覆盖)。
功能
| # | 功能 | 说明 |
|---|---|---|
| 1 | 🎤 语音输入 | 输入框旁麦克风按钮:点一下开始说话,识别结果逐句实时进输入框(中间结果实时显示在上方气泡);聆听中随时可打字改错字、删字——语音继续实时追加,删掉的内容停止后也不会回填 |
| 2 | 🔊 回复朗读 | 每条助手回复旁小喇叭,一键朗读该条;点击变「停止」随时打断 |
| 3 | 🔁 自动朗读 | 设置页开启后,新回复完成自动朗读(可随时打断) |
| 4 | ⚙️ 设置页 | dsh 设置 → ChatVoice:识别语言 / 自动朗读 / 音色 / 语速,保存即生效,无需重启 |
| 5 | 🛡 错误提示 | 麦克风权限被拒 / 浏览器不支持 / 非安全上下文 / 识别网络失败,全部有可读 toast,绝不静默失败 |
| 6 | 🇨🇳 中文优先 | zh-CN 识别 + 自动选择 Edge 内置 Xiaoxiao Online (Natural) 免费中文自然音色 |
为什么推荐 Edge
| 能力 | Chrome | Edge | 说明 |
|---|---|---|---|
| 语音识别 | ✅(识别走 Google 服务器) | ✅(识别走 Azure,国内更稳) | 国内网络下 Chrome 可能报 network 错误 |
| 朗读音色 | 部分在线音色 | ✅ Xiaoxiao Online (Natural) 免费中文最自然 | 在线音色需联网 |
| 麦克风(安全上下文) | 仅 localhost/HTTPS | 同左 | dsh web 默认 http://127.0.0.1:3080 ✅;LAN IP 访问麦克风不可用(朗读不受影响) |
安装
dsh plugin --profile web add dsh-chatvoice
# 或手动: pnpm add dsh-chatvoice(dsh.profile.bundles 会自动 reconcile)
重启 dsh web(dsh web),打开 http://127.0.0.1:3080 即可。
⚠️ 必须用 127.0.0.1 访问:语音识别需要安全上下文(HTTPS 或 localhost),LAN IP 直连时麦克风会被浏览器禁用(自动禁用输入功能并提示,朗读仍可用)。
使用
- 语音输入:点输入框工具条上的 🎤 → 浏览器弹麦克风授权(允许)→ 说话(确认句逐句实时进输入框、中间结果实时显示在上方气泡)→ 再点 🎤 停止 → 回车发送;识别中随时可以打字改错字甚至全删——语音只往框尾追加、绝不回写,你删掉的内容停止后也不会复活
- 朗读:点助手回复旁 🔊 → 开始朗读(按钮变红色 ⏹)→ 再点停止
- 自动朗读:设置 → ChatVoice → 勾选「自动朗读新回复」→ 保存,立即生效
Showing the opening section of the README — the full document lives in the repository以上为 README 开头摘要,完整文档在仓库内 · View the full README on GitHub →在 GitHub 查看完整 README →
nexu-io/open-design
Devin-AXIS/iPolloWork
liustack/modlens
Alisa0808/vox-director
EthanYoQ/AI-Novel-Writer
Anionex/agent-vision-toolkit
ysr666/dsh-vision-router
tong-io/tongflow