Pixcli Creative Production Toolkit
Paste the following prompt into your AI chat to install this skill:
Please follow https://skillhub.cn/install/skillhub.md to install @user_15292d5a/yjkj-pixcli.
About this skill
The Problem
Creative agent workflows often split into many model-specific steps: image generation, editing, background removal, SVG vectorization, video generation, voiceover, music, sound effects, and podcast assembly. Calling these models directly forces the agent to choose providers, craft prompts, poll jobs, download files, and avoid common pitfalls such as face drift, long video timeouts, and inconsistent audio stitching. pixcli collapses that surface into one command-level entry point.
How It Works
- Images and assets:
imagehandles text-to-image, reference images, ratios, transparency, and batches, with--vectorizeforSVGlogos and icons. - Editing and vector output:
editcovers image edits, enhancement, background removal, and upscaling;vectorizeconverts raster graphics into scalableSVG. - Video and audio:
videosupports text-to-video, image-to-video, extension, transitions, and lipsync;voice,music,sfx, anddialoguecover narration, scoring, effects, and multi-speaker dialogue. - Podcasts:
podcastor thesipcastalias turns a topic into a finished show, including script, voices, music, cover art, and a shareable page. - Final assembly: generated assets can be assembled into a finished video with Remotion templates.
It behaves like a creative task router: the agent describes the desired result, while the CLI/API classifies the job, enriches the prompt, selects a model, and returns parseable output.
Limits and Notes
- It expects a
PIXCLI_API_KEYand a Node.js runtime. - Video jobs can run long; submission plus polling is safer than blocking execution.
- Identity-preserving edits should avoid drift-prone models; the docs recommend identity-preserving routes for people.
- Podcast requests should use the full podcast command rather than manually stitching audio assets.
Use Cases
- Generate icons, logos, and hero visuals for a product launch, then convert raster images to scalable SVG.
- Turn a topic into a multi-speaker podcast with script, voices, music, cover art, and a share page.
- Create product video clips from reference images, then assemble them into a finished video with templates.
- Change backgrounds, remove backgrounds, upscale assets, or create try-on visuals for campaign images.
Best For
- Brand operations staff who need to produce logos, icons, and product visuals in batches.
- Creators making AI podcasts or voiceover videos with scripts, voices, music, and final edits.
- Agent engineers who want one CLI for image, video, and audio model calls.
- E-commerce visual operators who need try-on, cutout, upscaling, and hero product images.
Related Skills
Uses step-by-step choices to confirm business, palette, and layout, then exports an editable .drawio architecture diagram.
An agent skill for the Miaoyin AI music REST API that supports SUNO/Mureka-based song generation, continuation, cover, video, WAV, and stem splitting.
Distills text, files, or images into structured knowledge and generates knowledge-card prompts for drawing tools, supporting 3:4 portrait or 16:9 landscape outputs.
Capture full-page, viewport, or element screenshots with Playwright while handling lazy loading, internal scroll, mobile layouts, and Linux CJK fonts.