What problem does it solve?
Turning a 5-second voiceover line or abstract concept into a visually coherent B-roll clip normally requires manual storyboarding, image generation, and video assembly. This Skill automates that pipeline while preventing wasted video-generation spend through mandatory human approval checkpoints.
Core Features & Use Cases
- Three-Gate Approval Workflow: Proposes visual metaphors first (Gate 1), generates color collage stills via siliconflow-img-gen/Seedream only after confirmation (Gate 2), then produces video via aigc-video-gen i2v first/last-frame interpolation (Gate 3).
- Consistent Editorial Style: Enforces a unified design language—flat bold color fields, black-and-white halftone cut-outs, colored cardstock accents, cream keylines, and assemble-from-empty motion—across batched clips.
- Automated QA and Delivery: Builds contact sheets with ffmpeg, compares video end frames against approved stills, runs the video-review technical gate, and delivers silent 9:16 720x1280 5-second MP4s by default.
- Use Case: A content creator pastes five voiceover lines for a short video; the Skill proposes one visual metaphor per line, generates approved collage stills, and outputs five ready-to-use vertical B-roll clips.
Quick Start
Ask the agent to turn this voiceover script into collage-style B-roll clips using the collage b-roll workflow, then approve the proposed visual metaphors and still frames at each gate.