dsh-xiapan-media
Run the following command in DeepSeek Harness:
dsh plugin install dongsheng123132/dsh-xiapan-media
Paste the following prompt into your AI chat to install this plugin:
To install the plugin in DeepSeek Harness, run the command dsh plugin install dongsheng123132/dsh-xiapan-media from https://github.com/dongsheng123132/dsh-xiapan-media.
About this plugin
DeepSeek Harness excels as a powerful text model inference framework, but it lacks native multimedia processing capabilities, often requiring users to rely on external tools for image and video tasks. The dsh-xiapan-media plugin elegantly solves this gap by integrating three core media capabilities: image recognition/OCR, which allows users to directly paste images in sessions, leveraging Xiapan Cloud's qwen3.7-flash model to transcribe them into text for further analysis; image generation/editing, powered by gpt-image-2, supporting the creation of 1-4 images with customizable size, quality, and reference inputs; and video generation, using the Seedance model to enable text-to-video and image-to-video conversions, with options ranging from 5-15 seconds and resolutions from 480p to 1080p. This plugin is ideal for developers, AI researchers, and content creators using DeepSeek Harness, particularly those working on complex projects that require a unified workflow for text, images, and videos, as it streamlines tasks without the need to switch to external services.
Use Cases
- Directly paste images in DeepSeek Harness sessions for recognition and analysis.
- Generate or edit images using the plugin for design, marketing, or content creation.
- Convert text descriptions or static images into videos for multimedia projects.
Best For
- Developers integrating multimedia processing capabilities into AI applications.
- Content creators needing quick generation of visual content like images and videos.
- AI researchers conducting experiments and analysis related to images and videos.
Related Plugins
ModLens is a vision plugin for DeepSeek Harness that gives text-only models sight by reading images pasted directly into chat, with zero-config setup and multiple vision engines.
On-demand vision for text-only DeepSeek Harness agents: built-in free keyless vision chain and 14 vision tools, routing image turns as tool calls to vision models with pixel fidelity, no Python needed, one-command install.
Give text-only models in DeepSeek Harness eyes, enabling image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks.
An armor-breaking plugin for DeepSeek: rewrites instruction priority, expands refusal-blocked phrases, persists across sessions, and shows a green active indicator.