ai-avatar-video

Route inputs to avatar-generation models via the RunComfy CLI.

5|2|Updated May 18, 2026
One-click install
npx skills add https://github.com/doany-ai/skills --skill ai-avatar-video-doany-ai
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: ai-avatar-video
Source: https://github.com/doany-ai/skills/tree/main/ai-avatar-video
Command: npx skills add https://github.com/doany-ai/skills --skill ai-avatar-video-doany-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill lets you create engaging AI avatar videos from a single portrait and audio or scripted prompts, automating the most common animation workflows without manual rigging.

Core Features & Use Cases

  • Multi-model routing: picks the right RunComfy model path (OmniHuman, Wan-2-7, Wan-2-2 Animate, HappyHorse, Seedance) based on inputs.
  • Versatile outputs: supports photoreal portraits, stylized characters, lip-sync to audio, and cinematic scenes.
  • End-to-end workflow: from portrait URL or image to final video with minimal commands via the RunComfy CLI.

Quick Start

Install the RunComfy CLI and run a sample avatar video with a portrait and audio using ai-avatar-video.

Frequently Asked Questions about ai-avatar-video

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create an AI avatar video from a portrait and audio file?

To create an AI avatar video, provide a portrait image and an audio file to trigger the OmniHuman model for lip-sync animation. The Skill routes your inputs automatically and uses the RunComfy CLI to generate and download the final video.

What is the best way to generate lip-sync animation from a single portrait?

The best way to generate lip-sync animation from a single portrait is using the Wan 2-7 open-weights model. This Skill routes your portrait and audio inputs to the appropriate model to automate the animation workflow without manual rigging.

Can I use this Skill to animate a stylized full-body character?

Yes, you can animate a stylized full-body character by routing your inputs to the Wan-2-2 Animate model. The Skill supports both photoreal portraits and stylized characters for versatile avatar video generation.

Do I need an audio file to generate a cinematic avatar scene?

You do not need an audio file to generate a cinematic avatar scene. You can use scripted text prompts with the Seedance v2 Pro model for cinematic scenes, or HappyHorse for script-to-video generation, bypassing the audio requirement.

How does the Skill choose which AI model to use for avatar video generation?

The Skill chooses the AI model based on your inputs: it checks whether an audio file is provided and whether the subject is photoreal or stylized. It then routes the request to OmniHuman, Wan 2-7, Wan-2-2 Animate, HappyHorse, or Seedance accordingly.

What do I need to install to run AI avatar video workflows locally?

You need to install the RunComfy CLI to run AI avatar video workflows. The Skill invokes the CLI with the appropriate model ID and payload, then downloads the generated avatar video results to your specified output directory.