Tmall & Taobao Main Image Video Maker
Paste the following prompt into your AI chat to install this skill:
Please refer to https://skillhub.cn/install/skillhub.md to install @beatra-ai/taobao-main-video-maker.
About this skill
The E-Commerce Video Production Problem
Tmall and Taobao's main product videos have strict specifications: they must be clear, concise, and focus exclusively on the product. The first frame must be recognizable, and the process must preserve all necessary merchant-provided information like packaging, labels, and colors. These are typically required to be silent. Creating a short video that meets these standards and highlights selling points manually from a single static product image is time-consuming and technically challenging for many sellers.
This skill is designed to solve precisely this problem. It automates the entire image-to-video workflow, with the core objective of generating a static showcase video that is strictly product-centric, information-complete, and cleanly styled, ideal for new listings and detail page updates.
How the Skill Works: Core Workflow & Technical Paths
Starting with a real product image provided by you, the skill follows a structured workflow:
- Product Story Planning: Before generation, it automatically plans a three-act structure: clear product entrance, showcasing a detail or usage moment, and a clean exit returning to the product. This ensures narrative completeness.
- Fidelity-First Path Selection: Based on your image state and final goal, it selects the most faithful technical path:
- Strict First-Frame Animation: If the product image needs no adjustment, it uses it directly as the animation start point, ensuring 100% preservation of product information in the first frame and all subsequent motion.
- First-Frame Transform then Animate: If canvas or composition adjustments are needed, it first uses
beatra.images.transformto generate a new product first frame, then animates from that. - Reference Video Generation: Can incorporate the dynamic style from a reference video, but the product image remains the primary visual guide.
- First & Last Frame Interpolation: Used when precise control over both the start and end frames is required.
- Silent Video Generation: The skill defaults to and emphasizes generating silent video (
generate_audio: false), completely avoiding speech or background audio to meet the pure showcase requirement of main videos. The final output strictly adheres to the real-time model card's resolution, duration, and canvas settings. - Rigorous Confirmation & Submission: Before all paid operations (like image transforms, video generation), a confirmation card is generated, explicitly listing source material order, selected tools, models, cost estimates, and requiring you to confirm sufficient credits before submission, preventing unexpected charges.
Applicability Boundaries & Important Notes
Use this skill in the following scenarios:
- Core Use Case: Creating silent, product-focused Taobao/Tmall main videos starting from real product photos.
- Best For: New product launches, seasonal product page updates, and products where showcasing appearance, color, packaging, or a simple usage moment is key.
Use other skills for these cases:
- Short-form sales videos featuring a person on camera or with voiceover →
product-video-studio - Static main product images or image sets →
ecommerce-listing-image-setorproduct-photo-studio - Product showcase videos for WeChat Channels →
wechat-channels-product-video - Editing or repurposing existing videos →
beatra-ai-video-studio
Critical Constraints:
* Execution relies on the bundled MCP client for all remote calls. Direct use of the REST API is prohibited.
* All generations use real-time model card parameters. Final canvas, resolution, and duration are determined by model capabilities and cannot be manually set beyond what the model allows.
* Strictly silent mode: No audio elements of any kind are attached. Core parameters like product image order and reference sequence, once frozen, constitute a new paid job if changed.
Use Cases
- A product operator needs to quickly create a concise, product-focused main video for a newly listed item to use as the primary image on its detail page.
- During a seasonal change, the need to update main videos for multiple core products in a store, requiring a unified style that accurately presents new colors and packaging.
- A product has special surface textures or color details that require a silent dynamic sequence to showcase these selling points, overcoming the limitations of static images.
- A need to produce a clean, distraction-free product showcase video for a best-selling item in a platform campaign, suitable for autoplay scenarios.
Best For
- E-commerce operators who need to batch-create product main videos for daily new arrivals and page maintenance.
- E-commerce visual designers who transform product photography into dynamic video assets, requiring strict preservation of the product's true appearance and information.
- Brand e-commerce managers responsible for maintaining visual consistency in official flagship stores, needing to quickly generate platform-compliant videos that highlight the product itself.
- E-commerce photographers who need to convert their product photos into short videos to present product details more vividly.
Related Skills
Restyle a short video into a new visual style while preserving core elements such as characters, actions, and composition, suitable for various creative conversions like anime, illustration, ink wash, etc.
Create Douyin vertical video covers from topics, hooks, or materials with support for creative generation, image synthesis, and refinement.
An AI tool that transforms real photos into specified illustration styles while preserving subject recognition.
An engineering-driven solution that integrates design styles, UX workflows, design systems, and multi-platform implementation to solve cross-project design consistency.