What problem does it solve?
Turning written scripts into spoken audio requires picking the right voice, locale, and accent, and naive TTS calls often produce mismatched accents or unusable output. This Skill routes text-to-speech requests through OpenSquilla's audio tools with locale-aware voice selection and preview-first generation.
Core Features & Use Cases
- Locale-aware voice selection: Searches for voices matching the target language, locale, and accent (e.g., zh-CN Mandarin, en-GB British) before calling the TTS tool.
- Preview-first workflow: Generates a short sample for voice approval before committing to long batch narration.
- Batch narration: Splits long scripts into natural paragraphs under provider limits and produces stable output filenames.
- Use Case: A creator has a short-video script with VOICEOVER lines in Mandarin. The Skill searches for a Mandarin-capable voice, generates a one-paragraph preview, and after approval produces the full playable audio artifact.
Quick Start
Generate a Mandarin voiceover audio file from my video script using a natural female narrator voice.