dsh-voice-input
Run the following command in DeepSeek Harness:
dsh plugin install lougyang/dsh-voice-input
Paste the following prompt into your AI chat to install this plugin:
Run dsh plugin install lougyang/dsh-voice-input in the DeepSeek Harness terminal to install the plugin; the source repository is https://github.com/lougyang/dsh-voice-input. After installation, restart the web profile to activate the voice engine.
About this plugin
Your keyboard is not the only way to type. dsh-voice-input turns voice input into a system-wide habit: press and hold a hotkey anywhere on your desktop, speak, release, and the transcribed text lands at your cursor — in a browser, a chat app, a word processor, a search box, anywhere text is accepted. No window switching, no app launching, the whole desktop becomes an input field.
Under the hood, a locally running Whisper small model handles recognition with solid Chinese accuracy, automatically using CUDA acceleration when an NVIDIA GPU is detected, delivering results in roughly one to two seconds. The model is lazily loaded, so it consumes no GPU or RAM until you actually speak. Beyond transcription, the plugin offers pure audio recording with WAV export, a searchable history of every clip and transcript, and a tray settings panel for adjusting hotkeys, toggling push-to-talk mode, selecting model size, language, inference device, and recording directory — changes take effect instantly and survive plugin updates.
A good fit for writers and knowledge workers who produce heavy text volume and want to trade keystrokes for words; for privacy-conscious users who would rather keep voice data on their own hardware than ship it to a cloud API; and for anyone who simply likes the idea of a small microphone icon in the system tray meaning: press, hold, and just talk it out.
Use Cases
- Voice-dictate a search query in any browser field with a single hotkey
- Dictate paragraphs into a document without lifting fingers from the keyboard
- Capture a quick audio memo by pressing and holding the record key
Best For
- Writers and knowledge workers with heavy daily text output
- Privacy-conscious users who prefer local over cloud processing
- Developers with NVIDIA GPUs seeking local accelerated inference
Related Plugins
A unified suite combining hot runtime injection, task-aware thinking-mode routing, and a graded session protocol with red-team gates to sustain model diligence across long-horizon inference.
ModLens is a vision plugin for DeepSeek Harness that gives text-only models sight by reading images pasted directly into chat, with zero-config setup and multiple vision engines.
On-demand vision for text-only DeepSeek Harness agents: built-in free keyless vision chain and 14 vision tools, routing image turns as tool calls to vision models with pixel fidelity, no Python needed, one-command install.
Give text-only models in DeepSeek Harness eyes, enabling image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks.