Alan2Z/dsh-speak 预览 preview

Alan2Z/dsh-speak

Plugin插件 Native原生 ⭐ 11 MIT Vision & Media视觉与多媒体

让你的AI harness开口说话——一套已验证的语音播报解决方案

Project Overview项目介绍

dsh-speak is a speech announcement plugin originally built for DeepSeek Harness, with a native modular architecture that allows adapters for other AI coding harness tools like Claude Code. To install and activate the plugin in DSH, users only need to add the package to their dsh.profile.bundles, and it will auto-register via the included cordis.patch.yml, no manual patching required. It uses system-native speech synthesis, preferring natural voices on Windows and falling back to stock options if needed, while on macOS it uses the built-in say command that supports Siri natural voices.

The plugin automatically watches DSH’s session event stream, filtering out reasoning content and tool calls to only announce the final assistant reply. It notifies users when the agent is waiting for human approval or has asked a user-facing question via the ask_user_question API. Users can click a dedicated button in the DSH UI to replay any final reply at any time, and toggle optional announcements for task completions, tool errors, and goal changes independently.

dsh-speak is released under the open-source MIT license. On Windows, it requires Windows 10 or later with a working PowerShell environment; Windows 10 and newer Windows 11 24H2/25H2 require the NaturalVoiceSAPIAdapter to access natural voices, while older Windows 11 builds have natural voices pre-installed. On macOS, no extra dependencies are needed, and it supports both Intel and Apple Silicon chips. Diagnostic logs are stored in the system temporary directory for troubleshooting.

dsh-speak 是一款为 DeepSeek Harness 开发的原生语音播报插件,架构上支持适配其他 AI 编码框架(如 Claude Code)。核心功能是调用系统语音合成,将 AI 编码助手的最终回复、用户审批请求、任务完成通知等内容朗读出来,提醒用户不必持续盯着屏幕等待任务结束。它优先调用 Windows 系统原生自然语音,macOS 也可直接使用系统自带语音能力。

插件会自动监听 DSH 的会话事件流,只播报最终回复,跳过推理过程和工具调用内容,自动合并多步消息。它支持点击重播最终回复,支持独立开关各类事件播报,自带可视化设置页面,所有配置都可在 DSH 网页端完成,无需手动编辑配置文件。还有全局总开关可一键静音所有播报。

本插件采用 MIT 许可证开源,Windows 需要 Win10 或以上系统,搭配 PowerShell 环境,Win10 和新版 Win11 需要额外安装 NaturalVoiceSAPIAdapter 来桥接自然语音;macOS 无需额外依赖,原生支持 Intel 和 Apple Silicon 芯片。日志默认存储在系统临时目录方便排查问题。

Pre-install check安装前体检Compatibility · Security兼容性 · 安全性 1 note1 项提示
  • 11 stars - an early-stage project星标 11,属于早期项目
DSH walks through these 9 checksDSH 会逐条核对这 9 项

Compatibility兼容性

  • DSH, Node, OS and profile requirementsDSH 版本 / Node 版本 / 操作系统 / profile 是否满足要求
  • External dependencies and runtimes (Electron / Python / Docker, ...)外部依赖与运行时(Electron / Python / Docker 等)是否齐备
  • Conflicts with installed plugins: command names, skill / tool names, ports, duplicate MCP registration与已装插件是否冲突:命令名、skill / tool 重名、端口占用、重复 MCP 注册

Security安全性

  • Repo matches the facts registered here; archived or abandoned?仓库是否与页面登记一致,是否归档或长期停更
  • Safety of preinstall / install / postinstall and install.sh / setup.ps1preinstall / install / postinstall 与 install.sh、setup.ps1 是否安全
  • curl|bash, download-then-execute, obfuscation, unrelated domains → stop immediatelycurl|bash、下载即执行、混淆代码、无关域名 → 立刻停止
  • Typosquatting or unmaintained packages among the new dependencies新增依赖里有没有 typosquatting 或无人维护的包
  • Requested permissions vs. what the feature actually needs申请了哪些权限、是否超出功能所需(filesystem / network / shell / clipboard)
  • Any sudo / admin requirement, plus uninstall and rollback是否要求 sudo / 管理员权限,以及卸载与回滚方式

Anything uncertain must be marked unknown with a note on how to confirm it. This site's signal screen is a static snapshot, not a security audit.拿不准的必须标「未知」并说明要我怎么确认。本站的信号筛查是静态快照,不能替代安全审计。

Or use CLI install (for developers)或使用命令行安装(适合开发者)

CLI Install命令行安装

dsh plugin --profile web add dsh-speak

把 Alan2Z/dsh-speak 加入你的 DSH 配置(web profile)即可启用。

READMEREADME

dsh-speak 🔊 — Voice announcements for AI coding harnesses

English · 中文

鲸鱼娘大喇叭

Awesome DSH Plugin

npm version

Let your agent tell you when a long task is done — no more staring at the screen.

dsh-speak reads the final assistant reply aloud through system speech synthesis — on Windows using natural voices (Windows 11 built-in, or [NaturalVoiceSAPIAdapter] on Windows 10) with graceful fallback to stock voices; on macOS using the built-in say (can follow a Siri natural voice). It was built for DeepSeek Harness and is structured so any harness can plug in.

Features

  • Automatic: DSH web plugin watches the session event stream and announces the final reply (skips reasoning/tool-call narration, merges multi-step messages).
  • Gets your attention: announces approval requests (hears "需要你的审批" when the agent is waiting on you) and questions the agent asks via ask_user_question.
  • Final-reply replay (1.7.0): every final reply (turn tail) has a 🔊 button in its action bar — click to replay that message, click again to stop, click another to switch. Speech execution stays fully owned by the DSH host (keeps speaking even with the browser closed).
  • Host speech queue (1.7.0): only one native speech process runs at a time; queued items continue automatically. A WebSocket syncs the live state (which message is speaking, queue length) to the UI.
  • Optional event announcements (1.6.0): turn end, command done, goal changes, tool errors, and todo updates can each be announced, toggled independently (off by default).
  • Visual configuration (1.7.0): a dedicated Settings → dsh-speak settings page — every option (master switch, automatic speech, Markdown cleaning, code blocks, event toggles, fixed prompt, …) is editable from the Web UI, no hand-edited YAML.
  • Master switch (1.6.0): silence everything with one toggle.
  • Bundle auto-registration (1.3.0): declare the package in dsh.profile.bundles and the plugin registers itself via the bundled cordis.patch.yml — no manual patch entry needed.
  • Best-effort: never throws, never blocks the harness, never breaks a session.
  • Natural voices: Windows prefers natural voices — Windows 11 built-in packs, or voices registered via NaturalVoiceSAPIAdapter on Windows 10 (e.g. Xiaoxiao); macOS uses the system reading voice (Siri natural voices on recent macOS). Both fall back to any installed voice.
  • Robust text cleaning: strips markdown/URLs/emoji that make speech synthesis fail silently, and guards the adapter's per-utterance character ceiling.
  • Portable engine: any process can speak with one line: Windows powershell -File speak.ps1 -Text "你好" / macOS ./speak.sh -t "你好".

Showing the opening section of the README — the full document lives in the repository以上为 README 开头摘要,完整文档在仓库内 · View the full README on GitHub →在 GitHub 查看完整 README →

← 上一个 Prev dsh-grok-tui 下一个 Next dsh-prompt-library →