What problem does it solve?
This Skill eliminates the friction of generating high-quality text-to-speech audio, removing the need to navigate clunky web-based TTS platforms or manually configure complex audio settings for natural-sounding speech output.
Core Features & Use Cases
- Multi-Model ElevenLabs Integration: Access expressive default
eleven_v3, stable eleven_multilingual_v2, and fast eleven_flash_v2_5 voice models to balance quality, speed, and stability for different use cases.
- Custom Delivery Control: Use built-in audio tags like
[whispers], [short pause], and [excited] to adjust tone, pacing, and emotion for context-aware, natural speech.
- Use Case Example: Quickly generate a voiceover for social media content, create an accessible audio version of a written article, or produce a custom-toned voice response for a chat interface.
Quick Start
Use the sag skill to generate a natural-sounding text-to-speech audio file of your input text with your preferred voice and delivery style.