AI Agent Hub
Back to plugins
dsh-llm-guardian preview

dsh-llm-guardian

Model Inference Updated 2026.08.25

Run the following command in DeepSeek Harness:

dsh plugin install ice-kele/dsh-llm-guardian

Paste the following prompt into your AI chat to install this plugin:

To install this plugin in DeepSeek Harness, run the command 'dsh plugin install ice-kele/dsh-llm-guardian' and refer to the full source code at https://github.com/ice-kele/dsh-llm-guardian.

About this plugin

When managing multiple LLM providers with DeepSeek Harness, users often struggle with opaque service health and difficult usage tracking, which can lead to unexpected downtime or budget overruns. The dsh-llm-guardian plugin is crafted to address these challenges by enhancing model management locally. Its core capabilities include real-time health checks, local token quota management, and a comprehensive usage dashboard. Health checks enable periodic or on-demand testing of provider endpoints, allowing users to spot issues early and maintain reliability; local counters track token usage precisely, with configurable quotas that block requests when limits are reached, helping to control costs effectively. Additionally, built-in query adapters support account balance checks for platforms like DeepSeek and Z.AI Coding Plans, while the dashboard aggregates session logs into 7, 30, or 90-day views, featuring model filtering, activity heatmaps, and trend analysis for intuitive insights. This plugin is ideal for active DeepSeek Harness users, particularly developers, researchers, or teams relying on LLMs for daily work. It streamlines multi-provider monitoring and ensures privacy through local data processing and security measures—such as not storing API keys or uploading data—making the experience more reliable, cost-effective, and secure.

Screenshots

Use Cases

  • Monitor service health of multiple LLM providers in real-time within DeepSeek Harness.
  • Set local token quota limits to avoid unexpected API call overspending.
  • View session logs via dashboard to analyze model usage trends and efficiency.

Best For

  • Active DeepSeek Harness users needing unified management of multiple providers.
  • Developers and data scientists relying on LLMs for daily development and research.
  • Teams or organizations aiming to finely control API usage costs and performance.