What problem does it solve?
Converting a still image into a compelling video is hard because the “right” model depends on the user’s goal, such as portrait animation, product-style motion, lip-synced voiceover, or multimodal scene composition.
Core Features & Use Cases
- Intent-routed i2v model selection: Automatically matches the user’s intent to the best available RunComfy image-to-video route and model (default portrait/product animation, custom-audio lip-sync, or multimodal image+video+audio).
- Bundled prompting patterns per model: Uses the model-specific prompting approach to improve visual quality and reduce wasted iterations on the wrong setup.
- Flexible input schemas for different workflows: Supports
image_url for stills, audio_url for lip-synced talking-head results, and video_url/audio_url arrays for reference-driven multimodal clips.
Quick Start
Use the image-to-video skill to animate the file URL portrait.jpg into a short portrait video by matching natural facial motion and stable identity.