dubbing-skills

Convert long-form text into structured JSON dubbing scripts for Qwen3-TTS.

13|1|Updated Jan 25, 2026
One-click install
npx skills add https://github.com/mu-zi-lee/qwen3-tts-skill --skill dubbing-skills
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: dubbing-skills
Source: https://github.com/mu-zi-lee/qwen3-tts-skill/tree/main/dubbing-skills
Command: npx skills add https://github.com/mu-zi-lee/qwen3-tts-skill --skill dubbing-skills

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

Manually converting long-form content such as articles, video scripts, and multi-character dialogues into TTS-ready audio requires tedious text splitting, character assignment, and tone tuning, which is time-consuming and inconsistent for large volumes of content.

Core Features & Use Cases

  • Intelligent Text Splitting: Automatically splits text into natural speech segments while preserving quotation, ellipsis, and numeric boundaries to avoid awkward breaks.
  • Multi-Role Dialogue Support: Recognizes character markers in dialogues, assigns default speakers for different languages, and supports custom voice, voice design, and voice clone TTS modes.
  • Structured Dubbing Output: Generates standardized JSON dubbing scripts with tone instructions and configuration for batch TTS generation, suitable for audiobook production, video dubbing, and multi-role voiceover projects.

Quick Start

Provide the long-form text, video script, or multi-character dialogue you want to convert to audio, and the AI will automatically generate a structured dubbing script with character assignments, tone instructions, and TTS mode settings ready for batch voice generation.

Frequently Asked Questions about dubbing-skills

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert long-form text into a TTS-ready dubbing script?

You can convert long-form text into a TTS-ready dubbing script by inputting your content into the AI, which intelligently splits text while preserving boundaries, assigns characters, and outputs structured JSON for Qwen3-TTS batch generation.

How does intelligent text splitting handle quotation and sentence boundaries for audiobook production?

Intelligent text splitting for audiobook production automatically segments text into natural speech chunks while preserving quotation, ellipsis, and numeric boundaries to prevent awkward audio breaks during multi-role voiceover creation.

Can I use voice clone and custom voice modes for multi-role dialogue support?

Yes, multi-role dialogue support recognizes character markers in dialogues and assigns default speakers across different languages, while fully supporting custom voice, voice design, and voice clone TTS modes for character differentiation.

Does Qwen3-TTS batch audio generation work with multi-character dialogues in Chinese, English, Japanese, and Korean?

Qwen3-TTS batch audio generation works with multi-character dialogues in Chinese, English, Japanese, and Korean by processing the standardized JSON dubbing scripts to produce audio tailored for audiobook and video dubbing projects.

What is the best way to automate script splitting for video dubbing?

The best way to automate script splitting for video dubbing is using an AI system that applies automatic character recognition and tone tuning to generate standardized JSON dubbing scripts for batch TTS generation.