digital-avatar

Create and animate digital avatars from text or photos, outputting avatar_id and video-ready assets.

168|13|Updated Mar 20, 2026
One-click install
npx skills add https://github.com/npc-live/clawfirm --skill digital-avatar
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: digital-avatar
Source: https://github.com/npc-live/clawfirm/tree/main/app/assets/skills/video-skills/digital-avatar
Command: npx skills add https://github.com/npc-live/clawfirm --skill digital-avatar

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

本技能通过生成和动画化数字人形象,简化内容创作与视频制作流程,降低高成本的人物建模与口播制作门槛。

Core Features & Use Cases

  • 数字人生成:根据文本描述或真人照片生成稳定的数字人形象。
  • 口播视频生成:在同一流程内输出可用于视频剪辑的口播素材。
  • 多后端支持:可在 Kling、Jimeng、HeyGen、D-ID、Synthesia 之间切换以匹配语言与预算。
  • 工作流整合:结合后端配置实现自动化批量生成与渲染。

Quick Start

给出形象描述或上传照片,系统将创建数字人并输出 avatar_id 和可用于视频的资源。

Frequently Asked Questions about digital-avatar

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I create a digital avatar from a photo for video production?

To create a digital avatar from a photo, you provide the image and configure your chosen backend API key. The system processes the photo to generate a stable digital avatar, outputting an avatar_id and video-ready assets for your production workflow.

Can I use HeyGen and Synthesia backends for multilingual talking-head videos?

Yes, you can switch between HeyGen and Synthesia backends to generate multilingual talking-head videos. Configuring the specific API keys for your chosen platform allows the system to match the language and budget requirements of your video production.

What is the process to generate a virtual character from a text description?

Generating a virtual character from a text description involves entering your prompt and selecting a supported backend like Kling or Jimeng. The system uses your text to create the digital avatar and outputs the corresponding avatar_id and video assets.

Do I need API keys to animate digital avatars using this multibackend approach?

Yes, API keys are required to access and animate digital avatars using this multibackend approach. You must configure the specific credentials for your chosen backend, such as D-ID or Synthesia, to successfully generate and render the talking-head video assets.

Which backends are supported for AI voice cloning and avatar generation?

Supported backends for AI voice cloning and avatar generation include Kling, Jimeng, HeyGen, D-ID, and Synthesia. You can choose between these platforms based on your specific multilingual needs, budget constraints, and brand alignment requirements.

Why use multibackend support when creating digital avatars for brand-aligned content?

Using multibackend support when creating digital avatars allows you to switch between platforms like HeyGen and Synthesia. This ensures you can match specific language requirements, optimize production costs, and maintain consistent brand-aligned video output across different regions.