What problem does it solve? Turning a written Xiaohongshu (RED) note into an engaging vertical video normally requires manual storyboarding, animation work, voiceover recording, and video editing. This Skill automates the entire pipeline from Markdown content to a finished 1080x1920 MP4 with synchronized narration and background music. ## Core Features & Use Cases - End-to-end video pipeline: Analyzes a Markdown note, designs an 8-13 slide storyboard, builds GSAP-animated HTML slides with an industrial-neon design system, and renders a vertical MP4 via HyperFrames. - Deterministic TTS voiceover: Generates per-slide WAV narration with ChatTTS using a fixed seed (42) for reproducible results, then extracts durations into timing.json for precise audio-slide synchronization; Edge TTS is supported as a fallback. - Audio mixing and quality gates: Integrates per-slide voice clips plus looping BGM at volume 0.15 from a licensed Kevin MacLeod library, enforces lint checks, ffprobe verification, and a 20-point quality checklist. - Use Case: A content creator pastes a Markdown note about workplace management, and the Skill produces a 90-second kinetic typography video with Chinese voiceover, glass-morphism slide design, and background music ready for Xiaohongshu upload. ## Quick Start Use the xhs-video-workflow skill to turn my Markdown note into a vertical kinetic typography video with ChatTTS voiceover and background music.