Zhipu AI Image and Video Generation Pro
Paste the following prompt into your AI chat to install this skill:
Follow https://skillhub.cn/install/skillhub.md and install @user_3cce866c/glm-image-gen-pro into your AI assistant.
About this skill
Problem
Engineers building design assets, product demos, or short-form video often need to turn a text prompt into usable images, or animate a still image into a 6-second clip. Zhipu AI Image and Video Generation Pro wraps the Zhipu AI open platform APIs into a local skill, reducing manual request assembly, polling, and file download steps.
How It Works
- Text-to-image: supports
glm-image,cogview-4, andcogview-3-flash, with sizes such as1280x1280and quality set tohdorstandard. - Image-to-video: uses
CogVideoX-2to generate about 6 seconds of video from a remote image URL or a localJPG,PNG, orWebPfile. - Watermark control: after signing the disclaimer and completing real-name verification on the Zhipu open platform,
--no-watermarkcan be used. - Prompts: specify subject, style, lighting, camera, or motion, such as a British Shorthair cat with green eyes, a slow push-in shot, or soft backlighting.
Boundaries
Set ZHIPU_API_KEY before use. Outputs may be blocked by content review, video generation often takes 60–90 seconds, and long jobs may need --timeout. Save local files and temporary URLs promptly.
Use Cases
- Create e-commerce detail pages by generating 1280x1280 product hero images and style references.
- Prepare launch event shorts by converting a local PNG poster into a 6-second CogVideoX-2 preview.
- Deliver reviewed client assets using --no-watermark for clean watermark-free images.
- Draft short-video shots with prompts for slow push-in, pan, and orbit camera movement.
Best For
- E-commerce operations owners who need quick hero images, backgrounds, and style references.
- Solutions consultants who need to turn one static architecture image into a 6-second demo clip.
- Design leads who need to deliver watermark-free assets after content review.
- Application engineers who need to script text-to-image and image-to-video calls via API parameters.
Related Skills
A ComfyUI image-generation skill that uses a five-step dialog to collect prompts, prefill templates, confirm parameters, submit API jobs, and return results.
A design-system guidance skill for Impeccable that produces token-based rules, component states, accessibility criteria, and QA checklists.
Director-level video lapian that diagnoses material precision, then produces frame evidence, director analysis, a style bible, and a showcase video.
Plan and brief Amazon MAIN, Listing, and A+ image sets from verified product facts, then return plan and image QA status.