dsh-guide-dog
Run the following command in DeepSeek Harness:
dsh plugin install AtropinolTT/dsh-guide-dog
Paste the following prompt into your AI chat to install this plugin:
In DeepSeek Harness, install this plugin by running the command 'dsh plugin install AtropinolTT/dsh-guide-dog' from the source repository at https://github.com/AtropinolTT/dsh-guide-dog.
About this plugin
DeepSeek Harness is a powerful inference framework, but it lacks native visual input, limiting its ability to process images directly. The Guide Dog plugin bridges this gap by integrating MiniMax's multimodal services, giving DeepSeek 'eyes' and 'hands' to describe images, review designs, generate media, and interact via voice. Its core capabilities include VLM-based image description and structured inspection, multi-type content generation (images, videos, speech, music, text), inline Web UI preview and playback, and real-time voice conversation modes. Through automatic invocation of visual check tools, it ensures models handle visual tasks correctly even when the base model cannot see images.
This plugin is ideal for developers using DeepSeek Harness for content creation, frontend development, or multimodal tasks. Whether you need to review screenshots, generate promotional materials, or interact with an assistant via voice, Guide Dog offers seamless integration. It also lets users customize voice mode, input devices, and other preferences through a settings page, while planned accessibility features will further expand its reach. In short, Guide Dog transforms DeepSeek Harness from a text-only inference tool into a versatile assistant that can see, hear, speak, and create.
Use Cases
- Review frontend designs or screenshots to ensure visual elements are correct.
- Generate images, videos, and speech content for creation or presentation.
- Engage in real-time voice conversations with the assistant for improved interaction.
Best For
- Developers using DeepSeek Harness for content creation.
- Designers needing visual review and multimedia generation.
- General users seeking enhanced interaction experiences.
Related Plugins
ModLens is a vision plugin for DeepSeek Harness that gives text-only models sight by reading images pasted directly into chat, with zero-config setup and multiple vision engines.
On-demand vision for text-only DeepSeek Harness agents: built-in free keyless vision chain and 14 vision tools, routing image turns as tool calls to vision models with pixel fidelity, no Python needed, one-command install.
Give text-only models in DeepSeek Harness eyes, enabling image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks.
An armor-breaking plugin for DeepSeek: rewrites instruction priority, expands refusal-blocked phrases, persists across sessions, and shows a green active indicator.