What problem does it solve? Producing presenter-style videos normally requires cameras, actors, and editing time. This Skill lets you generate AI avatar videos programmatically through HeyGen's v2 API, with exact control over which avatar speaks, what it says, and how each scene looks. ## Core Features & Use Cases - Precise Avatar and Voice Control: List and preview avatars, select voices, and use each avatar's default voice for natural lip sync. - Multi-Scene Video Generation: Build videos with different avatars, scripts, backgrounds, captions, and text overlays per scene via POST /v2/video/generate. - Photo Avatars and Avatar IV: Turn uploaded photos or AI-generated portraits into talking presenters, including transparent WebM output for compositing. - Use Case: Create a three-scene product demo where a chosen avatar reads your exact script against branded backgrounds, then poll the video status endpoint and download the finished MP4. ## Quick Start Ask the agent to list available HeyGen avatars, pick one with its default voice, and generate a 1080p video speaking your provided script.