What problem does it solve? Producing narration, sound effects, and background music for videos, podcasts, or games normally requires voice actors, recording equipment, and audio licensing. This Skill generates all of that audio programmatically through the ElevenLabs API, including synchronization with Remotion video compositions. ## Core Features & Use Cases - Text-to-Speech: Generate voiceovers with model selection (multilingual_v2, flash, turbo, v3), voice settings presets, SSML pause and pronunciation control, and voice cloning from audio samples. - Sound Effects & Music: Create sound effects up to 22 seconds from text descriptions and instrumental music from 10 seconds to 5 minutes with genre, mood, and instrument prompts. - Remotion Integration: Sync generated audio to video scenes using manifest.json durations, per-scene audio components, fade/delay patterns, and demo-video playback rate matching. - Use Case: Write a VOICEOVER-SCRIPT.md for a product demo video, generate per-scene narration MP3s with a timing manifest, then bind them to Remotion Series sequences so visuals automatically match voiceover length. ## Quick Start Generate a professional voiceover MP3 from my narration script using the ElevenLabs multilingual model and save it for my video project.