Vidu Image Generation
Paste the following prompt into your AI chat to install this skill:
Please install @vidu/vidu-image-generate-2 according to https://skillhub.cn/install/skillhub.md.
About this skill
Problem Solved
The Vidu Image Generate skill addresses the repetitive work behind image generation calls: choosing a text-to-image or reference-based model, setting resolution and aspect-ratio parameters, selecting the correct API domain, polling an async task, and extracting the final accessible image URL. It turns these steps into a reusable workflow, so the caller only needs to describe the desired image or provide a reference image for generation or editing.
How It Works
- Task classification: plain text prompts use text-to-image generation; reference images trigger reference-based generation or image editing logic.
- Model selection:
q3-fastis preferred by default for speed, quality, and special aspect ratios;q2-prosuits high-quality output,q2-fastsuits fast generation,viduq2supports reference generation and editing, andviduq1is limited to reference-based generation. - API invocation: it checks
VIDU_API_KEY, selectsapi.vidu.cnorapi.vidu.combased on user language, submits the generation task, and polls for status. - Result return: after completion, it reads
creations[0].url, stores images under{baseDir}/uploads/by default, and returns the link as a Markdown file reference.
Boundaries and Notes
- A valid Vidu API key is required, and generated image URLs are generally valid for 24 hours, so downloads or saves should be handled promptly.
- Reference image limits differ:
viduq2accepts 0-7 images, whileviduq1requires 1-7 images. - For image editing, set
aspect_ratiotoauto; oversized inputs, invalid keys, or failed tasks should be retried after reviewing the error message.
Use Cases
- Generate product images for an e-commerce page from text descriptions, then save the assets.
- Use a reference image to generate a poster with a matching style and return the image link.
- Use viduq2 to perform partial repaint or image outpainting and output the new image.
- Select the domestic API domain for Simplified Chinese requests and complete async task polling.
Best For
- E-commerce operations staff who need text prompts to become product images
- Designers who need reference images to keep poster styles consistent
- Backend engineers who need to call the Vidu image generation API
- Developers who need to write generated image links into Markdown reports
Related Skills
Uses step-by-step choices to confirm business, palette, and layout, then exports an editable .drawio architecture diagram.
An agent skill for the Miaoyin AI music REST API that supports SUNO/Mureka-based song generation, continuation, cover, video, WAV, and stem splitting.
Distills text, files, or images into structured knowledge and generates knowledge-card prompts for drawing tools, supporting 3:4 portrait or 16:9 landscape outputs.
Capture full-page, viewport, or element screenshots with Playwright while handling lazy loading, internal scroll, mobile layouts, and Linux CJK fonts.