avatar-video

Automate talking avatar video creation from script to lip-synced output.

5|Updated Mar 27, 2026
One-click install
npx skills add https://github.com/barkleesanders/claude-code-starter --skill avatar-video-barkleesanders
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: avatar-video
Source: https://github.com/barkleesanders/claude-code-starter/tree/main/skills/avatar-video
Command: npx skills add https://github.com/barkleesanders/claude-code-starter --skill avatar-video-barkleesanders

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires jq, bc, curl, nano-banana, fal.ai, and includes scripts (resource) components.

What problem does it solve?

This Skill eliminates the manual, disjointed work of creating talking avatar videos by chaining portrait generation, voice synthesis, and lip-sync animation into a single automated pipeline, saving hours of editing and tool switching.

Core Features & Use Cases

  • End-to-End Video Pipeline: Automatically chains text script → portrait image → voice audio → lip-synced video in one command.
  • Flexible Quality Tiers: Choose premium OmniHuman 1.5 for high-quality short-form content, or budget Infinite Talk for long-form, low-cost videos.
  • Customizable Inputs: Skip automated steps to use your own existing portrait images or pre-recorded audio files.
  • Use Case: Create a 30-second app onboarding video with a professional avatar narrator, or generate bulk training content with consistent presenters for your team.

Quick Start

Use the avatar-video skill to generate a 30-second premium talking avatar video from the script "Welcome to our new product update".

Frequently Asked Questions about avatar-video

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate talking avatar video generation from a text script?

Talking avatar video generation is automated by chaining text scripts, portrait image generation, voice synthesis, and lip-sync animation into a single workflow pipeline. This eliminates manual tool switching and produces a finished video output.

What is the best way to create bulk training videos with a consistent presenter?

Creating bulk training videos with a consistent presenter is best achieved using budget-friendly long-form lip-sync animation tiers. This approach generates multiple video assets efficiently while maintaining presenter consistency across all training materials.

Can I use my own pre-recorded audio files for lip-sync animation?

Yes, you can use your own pre-recorded audio files for lip-sync animation by skipping the automated text-to-speech synthesis step. The pipeline allows skipping automated steps to supply custom portrait images or audio.

Does this avatar video pipeline work with OmniHuman 1.5 for high-quality short-form content?

Yes, the avatar video pipeline integrates with OmniHuman 1.5 to produce high-quality short-form content. It also supports Infinite Talk for long-form, low-cost videos, allowing you to select configurable quality tiers based on your needs.

Do I need fal.ai and nano-banana to generate text-to-video avatar clips?

You need fal.ai and nano-banana dependencies to coordinate the text-to-video avatar generation pipeline. These services handle the underlying portrait generation, voice synthesis, and lip-sync animation processing required for the automated workflow.

When should I choose Infinite Talk over OmniHuman 1.5 for AI video generation?

You should choose Infinite Talk over OmniHuman 1.5 for AI video generation when producing long-form, low-cost videos. OmniHuman 1.5 is better suited for high-quality short-form content where visual fidelity is the priority over cost and duration.