AI Agent Hub
Back to skills
ADP Multimodal Generation icon

ADP Multimodal Generation

Design & Media Updated 2026.08.29

Paste the following prompt into your AI chat to install this skill:

Please follow https://skillhub.cn/install/skillhub.md to install @user_33aa5c8b/search-baidusearch-baidusearch-baidusearch-baidusearch-baidusearch-baidusearch-baidusearch-baidusearch-baidusearch-baidu into your AI assistant.

About this skill

Problem

In Chinese business workflows, users may ask for image, video, or 3D model generation, while the backend may expose several ADP ability modules and model versions. If callers use raw identifiers or silently downgrade, requests can hit parameter mismatches, misclassified platform limits, or cross-modality substitutions. This skill normalizes Chinese intent into verified ADP multimodal abilities and executes them as asynchronous jobs.

How It Works

The supported modules are image.hy and image.gi for image generation, video.kling and video.vidu for video generation, and 3d.hy for 3D generation. Routing chooses the default ability: images default to image.hy, switch to image.gi when Gemini, nano banana, or another image model is requested; videos default to video.kling, switch to video.vidu when VIDU or another video model is requested; 3D remains 3d.hy. Only module aliases are accepted, not raw backend identifiers.

The flow has four steps: select the ability from user intent, choose a mode such as text2image, image2image, text2video, image2video, text23d, or image23d, execute the job through paired submit/query calls, and report the final ability, whether it was the default or keyword-triggered fallback, the failure stage if any, and platform limitations. It requires ADP_API_KEY and depends on requests.

Boundaries

This skill targets ADP platform multimodal abilities, not Tencent Cloud open APIs. Only fields verified in references/ability-catalog.md should be treated as confirmed facts. If an ability cannot be executed, report the blocker directly instead of silently downgrading across modalities, guessing internal interfaces, or swapping models without approval.

Use Cases

  • For a Chinese marketing page, turn product descriptions into candidate images and query the async result through `image.hy` or `image.gi`.
  • Turn a static first frame into a short video, routing to `video.kling` or `video.vidu` for `image2video` based on VIDU or Kling intent.
  • When delivering a 3D concept model, call `3d.hy` with `text23d` or `image23d`, then submit the async job and poll the result.
  • During multimodal integration tests, use only verified module aliases and distinguish whether a failure occurred at `submit` or `query`.

Best For

  • Backend engineers building Chinese multimodal content generation: route user intent to ADP image, video, or 3D modules and query results consistently.
  • Agent application developers integrating ADP: call verified `image`, `video`, and `3d` abilities without hard-coding low-level model IDs.
  • Platform engineers testing multimodal capabilities: identify whether failures occur at `submit` or `query`, and avoid silent cross-modality fallback.
  • AI product engineers for Chinese business workflows: switch image or video models based on keywords such as Gemini, nano banana, or VIDU.