voiceover

Generate synchronized Chinese and English video narration with edge-tts.

39|7|Updated Jan 23, 2026
One-click install
npx skills add https://github.com/MatrixReligio/ProductVideoCreator --skill voiceover
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: voiceover
Source: https://github.com/MatrixReligio/ProductVideoCreator/tree/main/.claude/skills/voiceover
Command: npx skills add https://github.com/MatrixReligio/ProductVideoCreator --skill voiceover

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill generates multilingual voiceovers for videos using edge-tts, enabling synchronized narration on a timeline and supporting adjustable speech rate and multiple voices.

Core Features & Use Cases

  • Multilingual narration: Chinese and English voices with diverse styles for product demos, tutorials, and marketing videos.
  • Timeline synchronization & validation: Segment-based voice generation with automatic duration checks to avoid overlaps or gaps.
  • Export & integration: Produces audio assets and metadata for downstream video composition or subtitle generation.

Quick Start

Provide the narration script or timeline file, select language and voice, and specify a timeline to generate and align the voiceover to the video.

Frequently Asked Questions about voiceover

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate synchronized voiceovers for videos in multiple languages?

Multilingual voiceovers synchronize narration to video timelines using edge-tts. Provide your script or timeline, select Chinese or English voices, and the Skill automatically generates audio segments, validates timing to prevent overlaps or gaps, and aligns them to your video.

Can I adjust speech rate and select different voices for voiceover narration?

Yes. The Skill supports adjustable speech rate and multiple voice options within Chinese and English languages, letting you customize narration tone and pacing for product demos, tutorials, and marketing videos.

What video production workflows benefit from timeline-based voiceover generation?

Timeline-based generation suits workflows requiring segment-by-segment narration with automatic duration validation. It prevents timing conflicts and produces audio assets and metadata ready for downstream video composition or subtitle generation.

Does voiceover generation work with edge-tts for both Chinese and English?

Edge-tts powers multilingual narration in Chinese and English with diverse voice styles. The Skill handles language-specific voice selection and delivers synchronized audio for immediate video integration.

What happens if voiceover segments overlap or have gaps on the timeline?

The Skill includes automatic validation to detect and flag overlaps and gaps before export. This prevents timing conflicts and ensures clean, continuous narration aligned precisely to your video segments.

Can I export voiceover audio and metadata for use in other video editing tools?

Yes. The Skill produces audio assets and metadata files designed for downstream integration with video composition tools or subtitle generation workflows, supporting your complete production pipeline.