lipsync

Route lip-sync requests across RunComfy endpoints and issue runcomfy run commands.

5|2|Updated May 18, 2026
One-click install
npx skills add https://github.com/doany-ai/skills --skill lipsync-doany-ai
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: lipsync
Source: https://github.com/doany-ai/skills/tree/main/lipsync
Command: npx skills add https://github.com/doany-ai/skills --skill lipsync-doany-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Lip-sync a face to an audio track across multiple RunComfy endpoints such as OmniHuman, Sync Labs, Kling, and Creatify.

Core Features & Use Cases

  • Route across multiple lip-sync endpoints to match user intent (avatar-style lipsync, mouth-swap on video, or script-driven generation).
  • Select the right model based on input shape, quality, and budget to deliver natural mouth motion.
  • Provide a consistent, repeatable invocation by using the documented runcomfy run payload.

Quick Start

Install the RunComfy CLI and invoke the appropriate model with your inputs to generate a lipsync result.

Frequently Asked Questions about lipsync

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I lip-sync a face to an audio track using RunComfy?

To lip-sync a face to an audio track, this Skill routes your inputs across RunComfy endpoints like OmniHuman, Sync Labs, Kling, and Creatify. It validates your inputs and issues the runcomfy run command with the correct model path and payload to generate natural mouth motion.

Can I generate avatar-style mouth movements from a still portrait?

Yes, you can generate avatar-style mouth movements from a still portrait. The Skill selects the appropriate lip-sync model based on your input shape and quality requirements, coordinating the endpoint selection to animate a static face with natural mouth motion synced to audio.

What is the best way to lip-sync an existing video to a new audio track?

The best way to lip-sync an existing video is to use the Skill to route your inputs to the appropriate RunComfy endpoint. It matches your intent for mouth-swap on video, validates the inputs, and executes the documented runcomfy run payload for consistent results.

Does this lipsync Skill support script-driven video generation?

Yes, this Skill supports generate-and-sync from a script. It coordinates endpoint selection across RunComfy APIs to generate video content from your script and applies lip-syncing to match the resulting avatar mouth movements to the generated audio.

Do I need the RunComfy CLI to perform lip-syncing?

Yes, you need the RunComfy CLI installed to perform lip-syncing. The Skill provides a consistent, repeatable invocation by using the documented runcomfy run payload, requiring the CLI to issue commands with the correct model path and input payload.