What problem does it solve?
This Skill eliminates the manual effort of working with a self-hosted Kokoro-FastAPI text-to-speech service, removing the need to independently parse API documentation, troubleshoot connection issues, or configure voice and pronunciation settings from scratch.
Core Features & Use Cases
- End-to-End TTS Workflows: Generate natural speech from text, discover available voices, and create custom blended voices for consistent audio output.
- Precise Pronunciation Control: Generate phonemes for exact pronunciation of names, slang, acronyms, and brand terms, with inline override support for tricky words.
- Use Case: A developer building a notification system can use this Skill to integrate the self-hosted TTS service, generate audio alerts with a custom voice blend, and resolve API errors without manual documentation lookup.
Quick Start
Use this Skill to generate a speech audio file from your input text using the default British male voice, or troubleshoot connection and permission errors with the self-hosted Kokoro TTS service.