seedance-v2

Generate cinematic lip-synced videos from text and reference media via RunComfy.

31|9|Updated Apr 30, 2026
One-click install
npx skills add https://github.com/agentspace-so/runcomfy-agent-skills --skill seedance-v2-agentspace-so
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: seedance-v2
Source: https://github.com/agentspace-so/runcomfy-agent-skills/tree/main/seedance-v2
Command: npx skills add https://github.com/agentspace-so/runcomfy-agent-skills --skill seedance-v2-agentspace-so

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill helps you generate short, cinematic, lip-synced videos from prompts and optional reference media, avoiding the manual effort of writing separate scripts and coordinating media.

Core Features & Use Cases

  • Cinematic short-form video generation: produces 4–15s outputs using RunComfy's bytedance/seedance-v2/pro endpoint.
  • Native in-pass audio with lip-sync: enables synchronized speech/SFX/music using generate_audio and optional audio references.
  • Multimodal identity and scene control: supports up to 9 images, 3 videos, and 3 audio references to keep characters consistent while scenes evolve.
  • Model routing guidance: instructs when to choose Seedance 2.0 Pro versus HappyHorse 1.0 / Wan 2.7 / Kling instead.

Quick Start

Ask the agent to generate a 9:16 lip-synced Seedance 2.0 Pro video by describing the scene and speaking tone, optionally adding an image reference for the character.

Frequently Asked Questions about seedance-v2

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate cinematic short-form video with native lip-synced audio?

To generate cinematic short-form video with native lip-synced audio, provide a text prompt describing the scene and speaking tone, then enable the generate_audio parameter to produce synchronized speech, SFX, or music within the 4–15 second output.

Can I use image and video references to keep characters consistent across generated scenes?

You can use image and video references to keep characters consistent by supplying up to 9 image URLs and 3 short video clips, allowing multimodal identity control while the scene evolves dynamically throughout the generated content.

What's the best way to create multi-language spokesperson dialogue ads?

The best way to create multi-language spokesperson dialogue ads is to provide a prompt specifying the speaking tone alongside an image reference for the character, utilizing native lip-synced audio for brand-consistent storytelling.

Does Seedance 2.0 Pro support custom audio references for dialogue generation?

Seedance 2.0 Pro supports custom audio references for dialogue generation by accepting up to 3 short audio clips, enabling precise voice matching and synchronized speech output alongside the generated cinematic video.

When should I choose Seedance 2. Pro over other video generation models?

Choose Seedance 2.0 Pro for cinematic short-form content requiring native lip-synced audio and multimodal identity control, while other models like HappyHorse 1.0 or Wan 2.7 may suit different video generation needs better.

What are the resolution and duration limits for cinematic video generation?

Duration limits for cinematic video generation range from 4 to 15 seconds, with configurable aspect ratios and resolution settings, supporting up to 9 image inputs and 3 video inputs for reference-guided scene visualization.