What problem does it solve?
Users need a fast way to generate cinematic short-form videos with native lip-synced audio using Seedance 2.0 Pro, without manually stitching references, timing, and playback.
Core Features & Use Cases
- Cinematic video generation with native lip-sync: Produces 4–15s outputs with synchronized in-pass audio from prompts and audio references.
- Multimodal reference support (identity, scene, voice tone): Combines up to 9 images, up to 3 short video clips, and up to 3 short audio references to keep the result coherent.
- Routing guidance vs sibling models: Uses Seedance for multimodal cinematic lip-synced shorts, while advising when to switch to HappyHorse, Wan, Kling, or LTX.
Quick Start
Tell the skill to create a 9:16 spokesperson-style video using Seedance 2.0 Pro with a prompt describing the shot and dialogue tone, optionally including an image reference for identity.