What problem does it solve?
This Skill helps creators produce polished short-form cinematic videos without stitching together separate video, voice, lip-sync, and sound-design workflows. It preserves the identity and visual details of reference media while generating coherent motion, dialogue, and synchronized audio through RunComfy.
Core Features & Use Cases
- Multimodal Reference Generation: Combine up to nine images, three videos, and three audio references to guide characters, environments, movement, voice tone, and scene continuity.
- Native Synchronized Audio: Generate dialogue, natural lip-sync, sound effects, music, and ambient sound in the same video pass.
- Cinematic Creative Control: Specify camera shots, motion, lighting, aspect ratio, duration, resolution, and reproducible seeds for short-form video production.
- Model Routing Guidance: Choose Seedance 2.0 Pro for reference-driven cinematic clips, while routing motion editing, audio-driven lip-sync, ultra-fast iteration, or maximum blind-vote quality to more suitable sibling models.
- Use Case: Create an eight-second vertical spokesperson advertisement using a character image, a conversational script, warm cafe lighting, and synchronized spoken audio.
Quick Start
Ask the AI to generate a five-second cinematic video from your prompt with synchronized audio using Seedance 2.0 Pro on RunComfy.