visual-review
Run the following command in DeepSeek Harness:
dsh plugin install wang-bool/visual-review
Paste the following prompt into your AI chat to install this plugin:
To install the visual-review plugin in DeepSeek Harness, use the command dsh plugin install wang-bool/visual-review, or manually install from the full source repository at https://github.com/wang-bool/visual-review.
About this plugin
In AI chat interfaces, images are powerful carriers of information, yet standard text models often fail to directly perceive visual content. This limitation hinders seamless interaction in environments like the DeepSeek Harness (DSH) Web interface, where users may upload images expecting insightful analysis. The visual-review plugin addresses this gap by enhancing the DSH Web interface with dual capabilities. It renders user-pasted or uploaded images directly within chat bubbles, ensuring visual clarity. Moreover, through intelligent interception and annotation conversion, the plugin enables any text model to "see" images and invoke the built-in visual_review tool for interpretation, returning descriptive text in Chinese that covers elements like text, objects, and scenes. Featuring a dual-engine approach, it prioritizes cloud-based OpenAI-compatible APIs for zero local dependencies, while automatically falling back to a local Qwen3-VL-8B model when unconfigured, safeguarding data privacy. This plugin is ideal for DSH Web users, including AI developers, researchers, and enthusiasts who regularly handle image content. It integrates image capabilities seamlessly without requiring a model switch, making it suitable for tasks ranging from quick local file analysis to customized queries (e.g., extracting chart trends). Users prioritizing data security benefit from the local engine, as all processing occurs on-machine with no external network requests.
Use Cases
- User pastes image into chat, model auto-interprets and returns description
- Analyze image files from disk to extract text or content
- Custom queries, such as describing chart trends or identifying objects and scenes
Best For
- AI developers integrating visual features into chat interfaces
- Researchers conducting image-related experiments and data analysis
- Everyday users handling image content, wanting models to respond directly
Related Plugins
ModLens is a vision plugin for DeepSeek Harness that gives text-only models sight by reading images pasted directly into chat, with zero-config setup and multiple vision engines.
On-demand vision for text-only DeepSeek Harness agents: built-in free keyless vision chain and 14 vision tools, routing image turns as tool calls to vision models with pixel fidelity, no Python needed, one-command install.
Give text-only models in DeepSeek Harness eyes, enabling image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks.
An armor-breaking plugin for DeepSeek: rewrites instruction priority, expands refusal-blocked phrases, persists across sessions, and shows a green active indicator.