What problem does it solve? Adding captions to talking-head videos usually means flat subtitle overlays that ignore the scene. This Skill composites captions into the footage itself — behind the subject via human matting, matched to scene lighting and palette — while keeping the source video untouched. ## Core Features & Use Cases - Identity catalog routing: Pick one of 32 visual identities (cream, ink, anchor, terminal, neon, glitch, etc.) from CATALOG.md; the engine, compiler, and authoring file are derived automatically. - Automated pipeline: One prepare script runs background-removal matting, WhisperX transcription, and safe-zone scene analysis in parallel, then compiles a small authored JSON into the final render. - Rail + embed caption model: A verbatim lower-third rail carries most text while scarce hero words are composited behind the subject with matte occlusion. - Use Case: Given a 30-second founder update clip, probe the scene, pick the keynote identity, author a small JSON, preview composite frames, and render a final.mp4 with word-timed captions embedded in the scene. ## Quick Start Add embedded captions to my talking-head video clip.mp4 using the most fitting identity from the catalog.