dsh-voice-input
forrestahha/dsh-voice-input
Voice-to-text input plugin for the DeepSeek Harness Web UI
3
stars
0
forks
MIT
License
2026-08-14
Created
2026-08-14
Last push
README
dsh-voice-input
English | 简体中文
Voice-to-text input for the DeepSeek Harness Web UI.
The plugin adds a microphone button immediately before the composer send button. It uses the browser's Web Speech API, streams recognition results into the current draft, and never submits the message automatically.
Features
- Adds to the official
conversation.input.rightextension slot. - Supports standard and WebKit-prefixed SpeechRecognition implementations.
- Uses the browser language, falling back to
zh-CN. - Preserves natural spacing for Chinese, Japanese, Korean, and Latin text.
- Stops recording when clicked again and aborts when the component unloads.
- Requires no additional API key or host-side service.
Install
Install directly from GitHub into the Web profile:
dsh plugin --profile web add github:forrestahha/dsh-voice-input
dsh web
For a pinned installation:
dsh plugin --profile web add github:forrestahha/dsh-voice-input#v0.1.1
Open the Web UI, select a workspace, and click the microphone button. The browser asks for microphone access on first use.
Browser support
The plugin requires SpeechRecognition or webkitSpeechRecognition. Chromium-based browsers provide the broadest support. Unsupported browsers show a disabled microphone button instead of failing at startup.
localhost is treated as a secure context by modern browsers. If Harness is served from another machine, use HTTPS or the browser may refuse microphone access.
Privacy
The plugin does not store audio and does not add its own network calls. The Web Speech API implementation is controlled by the browser and may send audio to the browser vendor's speech service. Review your browser's privacy policy before recording sensitive material.
Development
Requirements: Node.js ^22.19.0 || >=24.0.0 and pnpm 11.
pnpm install
pnpm check
Install a local checkout for integration testing:
dsh plugin --profile web add /absolute/path/to/dsh-voice-input
dsh web
Design
The npm package is both a Harness bundle and a client plugin:
cordis.patch.ymlinserts the package into the selected profile.- The Node entry is intentionally empty;
dsh.clientdiscovers./client. - The browser entry registers
VoiceInputButtoninconversation.input.right. - Harness supplies the current input state and
inputActions.setDraft()through slot props.
No agent-loop, model, session-log, or Host API behavior is changed.
License
MIT
More in Voice & Speech
dsh-voice-ai-girlfriend
by beiyege-01
语音 AI 女友(Voice AI girlfriend for DeepSeek Harness):Whisper 语音输入 + Qwen3-TTS 声音克隆 + 句子级流式朗读 + 数字人动画窗。插话/排队双模式,说话即打断。
dsh-voice-mic
by zachary7456
DeepSeek Harness (dsh) 语音输入插件:麦克风按钮/快捷键录音,实时转写回填输入框。三种识别引擎:浏览器 Web Speech、本地 SenseVoice/Paraformer 离线后端(一键部署)、OpenAI 兼容云端 ASR API。
deepseek-harness-voice-context
by charlesliuzc
DeepSeek Harness with Voice Context speech-to-text integration
dsh-voice
by stardustlc666
Voice pair: free edge-tts neural speech synthesis + OpenAI-compatible ASR transcription.
