What problem does it solve? Editing video normally requires timeline software, manual scrubbing, and preset-driven workflows. This Skill turns video editing into a conversation: it transcribes footage with word-level timestamps, proposes a cut strategy for your approval, then executes cuts, color grades, overlay animations, and subtitles through ffmpeg with production-correctness rules that prevent silent failures like misaligned captions or audio pops. ## Core Features & Use Cases - Transcript-driven cutting: Word-level verbatim ASR (ElevenLabs Scribe) produces phrase-level packed transcripts; cuts snap to word boundaries with padded edges and 30ms audio fades. - Full post-production pipeline: Per-segment extraction with lossless concat, ASC CDL-style color grading, burned subtitles applied last, and overlay animations built with HyperFrames, Remotion, Manim, or PIL. - Self-verifying renders: The Skill inspects its own rendered output at every cut boundary for visual discontinuities, audio pops, and subtitle occlusion before showing you a preview. - Use Case: You have five takes of a product launch talking-head video. The Skill transcribes all takes, picks the best take per beat (hook, problem, solution, CTA), builds animated overlay cards synced to narration, grades the footage, burns subtitles, and delivers a final 1080p cut. ## Quick Start Ask the assistant to edit the videos in your footage folder into a two-minute cut with subtitles and animated overlays, then approve the proposed strategy before rendering begins.