What problem does it solve? Creating cinematic short-form video with synchronized audio and lip-sync normally requires stitching together separate generation, TTS, and editing tools. This Skill routes video generation requests to ByteDance Seedance 2.0 Pro on the RunComfy Model API, handling multi-modal references (images, videos, audio) in a single call. ## Core Features & Use Cases - Multi-modal video generation: Combine up to 9 image references, 3 video clips, and 3 audio references with a text prompt to produce 4–15 second videos at 480p or 720p. - Native lip-synced audio: Generate in-pass synchronized speech, SFX, and music with tone direction, ideal for spokesperson ads and dialogue content. - Model routing guidance: Built-in decision table for when to use Seedance 2.0 Pro versus HappyHorse 1.0, Wan 2.7, Kling, or LTX 2. - Use Case: A marketer needs a vertical 9:16 lip-synced ad where a barista explains today's special. Provide a headshot image reference plus a prompt describing camera movement and tone, and the Skill invokes runcomfy run bytedance/seedance-v2/pro to produce the clip. ## Quick Start Ask the AI to generate a video with Seedance 2.0 Pro, for example: "Use Seedance to generate an 8-second 9:16 video of a barista explaining today's special with natural lip-sync, using this headshot as the image reference."