FuzzySoul/dsh-chatvoice 预览 preview

FuzzySoul/dsh-chatvoice

Plugin插件 Native原生 ⭐ 4 MIT Vision & Media视觉与多媒体

ChatVoice — DeepSeek Harness (dsh) 的免费语音输入与 AI 回复朗读插件。零配置/零成本/免 API key 语音插件。

Project Overview项目介绍

This is a native plugin built exclusively for the DeepSeek Harness web client, adding full voice functionality including speech-to-text for prompts and text-to-speech for AI responses. It relies entirely on the browser’s native Web Speech API, so no third-party API keys are required, there is no cost to use it, and you don’t need to do any extra configuration after installation. To install the plugin, run the command dsh plugin --profile web add dsh-chatvoice in your terminal, then restart the DSH web service and navigate to the local address 127.0.0.1:3080 to get started.

The plugin adds a microphone button next to the input box for voice input. Once you click the button and grant mic permission to your browser, confirmed speech results are added to the input box in real time, while intermediate results show in a temporary bubble at the top of the interface. You can manually edit or delete text in the input box while recognition is still running, and the plugin only appends new speech content instead of overwriting your edits.

You can customize the recognition language, auto-read toggle, voice, and speech rate in the DSH settings menu under the ChatVoice section, and changes take effect immediately without requiring a restart of DSH. For users in mainland China, the Edge browser is recommended, as it offers more stable speech recognition via Azure and access to the free natural-sounding Xiaoxiao Online Mandarin voice. Firefox and Safari do not support the SpeechRecognition API required for voice input, so only the text-to-speech feature works on these browsers.

这是专为 DeepSeek Harness (DSH) 开发的原生插件,为 DSH Web 版本添加了语音输入和 AI 回复文本朗读的完整语音功能闭环。它全程调用浏览器原生的 Web Speech API,无需申请第三方 API 密钥,零配置开箱即用,也没有任何额外成本。安装时只需要在命令行执行 dsh plugin --profile web add dsh-chatvoice,重启 DSH Web 服务后访问本地地址即可使用。

核心功能包含语音输入和回复朗读两大块,语音输入会在输入框旁添加麦克风按钮,点击开始识别后,确认内容逐句实时进入输入框,中间结果显示在顶部气泡,识别过程中随时可以手动编辑修改内容,语音只会追加不会回写已有内容。每条 AI 回复旁都有朗读按钮,点击即可一键朗读,随时点击打断,还支持开启新回复自动朗读功能。

插件支持自定义识别语言、朗读音色、语速,设置保存后立即生效无需重启。国内网络环境下更推荐使用 Edge 浏览器,语音识别走 Azure 服务比 Chrome 更稳定,还能获得自然流畅的免费中文在线音色。受浏览器 API 限制,Firefox 和 Safari 不支持语音识别,只能使用朗读功能,访问必须用 127.0.0.1,LAN IP 访问会被禁用麦克风。

Pre-install check安装前体检Compatibility · Security兼容性 · 安全性 1 warning1 项注意
  • Only 4 stars - very few users, little community feedback星标只有 4,几乎没人在用,遇到问题缺少社区反馈
DSH walks through these 9 checksDSH 会逐条核对这 9 项

Compatibility兼容性

  • DSH, Node, OS and profile requirementsDSH 版本 / Node 版本 / 操作系统 / profile 是否满足要求
  • External dependencies and runtimes (Electron / Python / Docker, ...)外部依赖与运行时(Electron / Python / Docker 等)是否齐备
  • Conflicts with installed plugins: command names, skill / tool names, ports, duplicate MCP registration与已装插件是否冲突:命令名、skill / tool 重名、端口占用、重复 MCP 注册

Security安全性

  • Repo matches the facts registered here; archived or abandoned?仓库是否与页面登记一致,是否归档或长期停更
  • Safety of preinstall / install / postinstall and install.sh / setup.ps1preinstall / install / postinstall 与 install.sh、setup.ps1 是否安全
  • curl|bash, download-then-execute, obfuscation, unrelated domains → stop immediatelycurl|bash、下载即执行、混淆代码、无关域名 → 立刻停止
  • Typosquatting or unmaintained packages among the new dependencies新增依赖里有没有 typosquatting 或无人维护的包
  • Requested permissions vs. what the feature actually needs申请了哪些权限、是否超出功能所需(filesystem / network / shell / clipboard)
  • Any sudo / admin requirement, plus uninstall and rollback是否要求 sudo / 管理员权限,以及卸载与回滚方式

Anything uncertain must be marked unknown with a note on how to confirm it. This site's signal screen is a static snapshot, not a security audit.拿不准的必须标「未知」并说明要我怎么确认。本站的信号筛查是静态快照,不能替代安全审计。

Or use CLI install (for developers)或使用命令行安装(适合开发者)

CLI Install命令行安装

dsh plugin --profile web add dsh-chatvoice

把 FuzzySoul/dsh-chatvoice 加入你的 DSH 配置(web profile)即可启用。

READMEREADME

ChatVoice 🎤🔊 — dsh-chatvoice

English | 中文

给 DeepSeek Harness (dsh) 装上「免费、免 API key、开箱即用」的语音输入 + AI 回复朗读闭环。 全程浏览器原生 Web Speech API —— 零配置、零成本、无任何后端与注册。

语音输入
🎤 语音输入:确认句逐句实时入框,中间结果进气泡

回复朗读
🔊 回复朗读:点小喇叭一键朗读,可随时打断

边听边改
✏️ 边听边改:聆听中打字修改/全删,语音只追加不回写

零配置 零成本 免 API Key npm MIT

ChatVoice = Chat + Voice:一个插件解决「嘴」和「耳朵」——写代码时手不离键盘,用嘴问 AI;懒得看长回复,让 AI 读给你听(听力型学习 / 无障碍 / 摸鱼躺用场景全覆盖)。

功能

# 功能 说明
1 🎤 语音输入 输入框旁麦克风按钮:点一下开始说话,识别结果逐句实时进输入框(中间结果实时显示在上方气泡);聆听中随时可打字改错字、删字——语音继续实时追加,删掉的内容停止后也不会回填
2 🔊 回复朗读 每条助手回复旁小喇叭,一键朗读该条;点击变「停止」随时打断
3 🔁 自动朗读 设置页开启后,新回复完成自动朗读(可随时打断)
4 ⚙️ 设置页 dsh 设置 → ChatVoice:识别语言 / 自动朗读 / 音色 / 语速,保存即生效,无需重启
5 🛡 错误提示 麦克风权限被拒 / 浏览器不支持 / 非安全上下文 / 识别网络失败,全部有可读 toast,绝不静默失败
6 🇨🇳 中文优先 zh-CN 识别 + 自动选择 Edge 内置 Xiaoxiao Online (Natural) 免费中文自然音色

为什么推荐 Edge

能力 Chrome Edge 说明
语音识别 ✅(识别走 Google 服务器) ✅(识别走 Azure,国内更稳) 国内网络下 Chrome 可能报 network 错误
朗读音色 部分在线音色 ✅ Xiaoxiao Online (Natural) 免费中文最自然 在线音色需联网
麦克风(安全上下文) 仅 localhost/HTTPS 同左 dsh web 默认 http://127.0.0.1:3080 ✅;LAN IP 访问麦克风不可用(朗读不受影响)

安装

dsh plugin --profile web add dsh-chatvoice
# 或手动: pnpm add dsh-chatvoice(dsh.profile.bundles 会自动 reconcile)

重启 dsh web(dsh web),打开 http://127.0.0.1:3080 即可。

⚠️ 必须用 127.0.0.1 访问:语音识别需要安全上下文(HTTPS 或 localhost),LAN IP 直连时麦克风会被浏览器禁用(自动禁用输入功能并提示,朗读仍可用)。

使用

  1. 语音输入:点输入框工具条上的 🎤 → 浏览器弹麦克风授权(允许)→ 说话(确认句逐句实时进输入框、中间结果实时显示在上方气泡)→ 再点 🎤 停止 → 回车发送;识别中随时可以打字改错字甚至全删——语音只往框尾追加、绝不回写,你删掉的内容停止后也不会复活
  2. 朗读:点助手回复旁 🔊 → 开始朗读(按钮变红色 ⏹)→ 再点停止
  3. 自动朗读:设置 → ChatVoice → 勾选「自动朗读新回复」→ 保存,立即生效

Showing the opening section of the README — the full document lives in the repository以上为 README 开头摘要,完整文档在仓库内 · View the full README on GitHub →在 GitHub 查看完整 README →

← 上一个 Prev dsh-websearch 下一个 Next dsh-store →