media-lipsync

Synchronize lips and drive talking-head animation from audio or video using LivePortrait and LatentSync models.

15|4|Updated Apr 18, 2026
One-click install
npx skills add https://github.com/damionrashford/media-os --skill media-lipsync
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: media-lipsync
Source: https://github.com/damionrashford/media-os/tree/main/skills/media-lipsync
Command: npx skills add https://github.com/damionrashford/media-os --skill media-lipsync

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

Open-source lip-sync and talking-head animation tools enable creators to generate realistically synchronized mouth movements and facial performance from audio or driving video without relying on proprietary APIs.

Core Features & Use Cases

  • Still-to-video lip-sync: animate a single image into a talking-head using audio or driving video.
  • Dubbing & ADR workflows: re-sync lips to new audio for language swaps.
  • License-safe workflows: coordinate models with permissive licenses and avoid restricted weights.

Quick Start

Run the lipsync driver to orchestrate LivePortrait or LatentSync from a still image or driving video to produce a talking-head synced to audio.

Frequently Asked Questions about media-lipsync

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I synchronize lips to new audio for dubbing workflows?

To synchronize lips for dubbing, this Skill uses LatentSync and LivePortrait to re-sync mouth movements in a clip to new audio. It coordinates permissive open-source models to handle language swaps without proprietary API dependencies.

Can I animate a still image into a talking-head using audio?

Yes, you can animate a single still image into a talking-head using audio or driving video. The Skill orchestrates LivePortrait and LatentSync workflows to generate realistic facial performance and mouth movements from the input media.

Does this lip-sync tool rely on proprietary APIs or restricted model weights?

No, this lip-sync tool avoids proprietary APIs and restricted weights entirely. It enforces permissive licenses for open-source models, ensuring end-to-end media creation remains license-safe for reliable production use.

What is the best way to drive expression retargeting from a video clip?

The best way to drive expression retargeting is using the CLI to orchestrate LivePortrait workflows. This maps facial performance from a driving video onto a still image or existing clip for accurate animation.

Are there guardrails for handling errors during talking-head animation processing?

Yes, the talking-head animation process includes built-in guardrails and error handling. These mechanisms orchestrate model workflows safely via the CLI, providing reliable production use during complex lip-sync operations.