dsh-model-modality
Run the following command in DeepSeek Harness:
dsh plugin install ct-jyjntc/dsh-model-modality
Paste the following prompt into your AI chat to install this plugin:
Run dsh plugin install ct-jyjntc/dsh-model-modality in DeepSeek Harness to install; source code at https://github.com/ct-jyjntc/dsh-model-modality
About this plugin
The DSH runtime enforces strict image-input gates: message images, the read_image tool, attachment uploads, and sub-agent image prompts all depend on the inputModalities field in model metadata. When using a third-party relay or a custom provider, that field is often missing or mis-declared, so a genuinely multimodal model gets blocked with Model does not support image input, image tools stay locked, and the underlying stream call silently projects images into text.
This plugin offers two conversational tools: declare_model_multimodal writes the image-support flag directly into the correct provider setting (inputModalities for llm-deepseek, input or modelOverrides for llm-pi-ai) and validates immediately via resolveModelInfo; list_model_multimodal gives an at-a-glance view of declared input modalities across providers. The declaration persists in the settings of the owning provider, independent of the plugin lifecycle, so uninstalling the plugin does not erase it.
Built for DSH users who work with relay stations, custom API endpoints, or third-party model catalogs. If the model itself supports image input but DSH has not unlocked the gate, a single tool call is enough to open it, with no source edits or manual JSON fiddling required.
Use Cases
- A third-party relay model supports images but is blocked by DSH's inputModalities gate
- The read_image tool needs to be unlocked so the agent can inspect local image files
- A custom API endpoint's model metadata lacks the modality field, causing the stream to silently flatten images into text
Best For
- DSH users who run relay stations or custom API endpoints
- Developers who need to backfill multimodal declarations in third-party model catalogs
- Power users who prefer a tool-call over manually editing JSON config to unlock image gates
Related Plugins
A unified suite combining hot runtime injection, task-aware thinking-mode routing, and a graded session protocol with red-team gates to sustain model diligence across long-horizon inference.
ModLens is a vision plugin for DeepSeek Harness that gives text-only models sight by reading images pasted directly into chat, with zero-config setup and multiple vision engines.
On-demand vision for text-only DeepSeek Harness agents: built-in free keyless vision chain and 14 vision tools, routing image turns as tool calls to vision models with pixel fidelity, no Python needed, one-command install.
Give text-only models in DeepSeek Harness eyes, enabling image Q&A, long-screenshot OCR, UI restoration, and GUI visual tasks.