What problem does it solve? Producing a professional explainer video normally requires separate tools for voiceover, imagery, music, animation, and video editing. This Skill orchestrates the entire pipeline from a text brief to a rendered MP4 using open-source AI models on cloud GPUs and Remotion for composition. ## Core Features & Use Cases - AI Asset Generation: Create per-scene voiceovers with Qwen3-TTS, background music with MusicGen, scene images with FLUX.2, b-roll clips with LTX-2, and talking-head narrators with SadTalker, all running on Modal cloud GPUs. - Remotion Composition: Assemble scenes, per-scene audio, narrator picture-in-picture, and transitions in React-based Remotion templates, then render to MP4. - Timing Synchronization: Measure actual voiceover durations with ffprobe and automatically adjust scene durations in the demo config. - Use Case: Write a product brief, generate a 60-second product demo video with narration, background music, and animated scenes for roughly $1-3 in cloud compute. ## Quick Start Ask the agent to create a 60-second product demo video for your product using the video toolkit, starting from a short text brief.