In-chat read-aloud for DeepSeek Harness: tap to read, pause and resume AI replies with natural Doubao TTS voices (BYOK), reading only the final answer with code, tables and diagrams filtered; local engine, plugin keeps no API key.
- GitHub stars
- 34
- License
- MIT
- Added
- 2026-08-28
GitHub info
- GitHub stars
- 34
- License
- MIT
- Primary language
- Swift
- Last push
- Aug 19, 2026, 2:56 AM
- Maintainer
- PolinniZhong
- Added
- 2026-08-28
Preview from README
Screenshots and GIFs extracted from the repository README (badges and avatars filtered out).

点读 → 暂停 → 继续 演示
https://raw.githubusercontent.com/PolinniZhong/dsh-omi-voice/main/docs/assets/reading.gif

开通豆包语音合成 1.0
https://raw.githubusercontent.com/PolinniZhong/dsh-omi-voice/main/docs/assets/doubao-2-enable.png

创建 API Key
https://raw.githubusercontent.com/PolinniZhong/dsh-omi-voice/main/docs/assets/doubao-3-create-key.png

找到豆包语音
https://raw.githubusercontent.com/PolinniZhong/dsh-omi-voice/main/docs/assets/doubao-1-find.png

未朗读状态
https://raw.githubusercontent.com/PolinniZhong/dsh-omi-voice/main/docs/assets/idle.png

朗读中状态
https://raw.githubusercontent.com/PolinniZhong/dsh-omi-voice/main/docs/assets/reading.png
Install
dsh plugin --profile web add dsh-omi-voiceTags
README badge
Add this Markdown to your plugin README to link back to its listing.
[](https://dshget.com/plugins/PolinniZhong/dsh-omi-voice)Related plugins
Voice & Audiodsh-speak
★ 8Zero-dependency, event-driven voice announcement plugin: no extra model, no token cost. Speaks with the system's built-in natural voice, supporting both Windows and macOS; final-reply announcements, approval & question alerts, optional event announcements (turn end, command done, goal change, tool errors, todo updates), replayable final replies, and a bilingual visual settings page.
dsh-gsv
★ 3Real-time local TTS for DeepSeek Harness: voice presets, auto-read, engine setup assistant, a read-aloud button, and a settings panel for the GSV-TTS-Lite engine.
dsh-talk
Voice-first session loop for DeepSeek Harness: a composer microphone button with browser/local speech-to-text (Web Speech, FunASR, whisper.cpp), a speak tool for text-to-speech replies (browser, edge-tts, piper), event announcements with mute, and speak-to-interrupt.