What problem does it solve?
Convert written text and scripts into natural-sounding spoken audio so users can listen to content instead of reading, enabling accessibility, voiceovers, and rapid content consumption.
Core Features & Use Cases
- Quick single-voice TTS for instant, low-latency MP3 streams useful for chat replies, notifications, and short reads.
- Scripted multi-speaker generation to produce dialogues, audiobooks, and multi-character voiceovers with per-segment speaker assignment.
- Configurable output modes and persistence including inline playback, local download, and saving default speaker preferences per language.
- Robust speaker selection and safety patterns: always fetch available speakers, respect API key presence, and follow interactive AskUserQuestion flows for enumerable choices.
Quick Start
Ask the skill to "朗读这段:" or "TTS this: The server will be down for maintenance at midnight."