sag

Generate speech from text via the ElevenLabs API with voice selection.

455|34|Updated Mar 9, 2026
One-click install
npx skills add https://github.com/understudy-ai/understudy --skill sag-understudy-ai
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sag
Source: https://github.com/understudy-ai/understudy/tree/main/skills/sag
Command: npx skills add https://github.com/understudy-ai/understudy --skill sag-understudy-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a command-line interface for generating speech from text using ElevenLabs, offering a user experience similar to the macOS say command.

Core Features & Use Cases

  • Text-to-Speech Generation: Convert written text into spoken audio using ElevenLabs voices.
  • Voice Selection & Customization: Choose from various ElevenLabs voices and apply specific pronunciation or delivery rules.
  • Use Case: Generate an audio file for a chatbot response in a specific character voice, like a "crazy scientist," for enhanced user engagement.

Quick Start

Use the sag skill to speak the phrase "Hello there" using the default voice.

Frequently Asked Questions about sag

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate speech from text using ElevenLabs TTS?

To generate speech from text using ElevenLabs TTS, you provide your text and an ElevenLabs API key, select a voice, and the Skill synthesizes an audio file for local playback. It functions similarly to the macOS say command.

Can I use ElevenLabs voice selection for a specific character like a crazy scientist?

Yes, you can use ElevenLabs voice selection to assign specific character voices, such as a crazy scientist, to your audio generation. The Skill supports applying delivery modifications and pronunciation rules to match the persona.

Do I need an ElevenLabs API key to use text-to-speech generation?

Yes, you need an ElevenLabs API key to authenticate and use the text-to-speech generation capabilities. The Skill relies on the ElevenLabs TTS API to synthesize spoken audio from your written text.

What is the best way to add voice responses to a chatbot using speech synthesis?

The best way to add voice responses to a chatbot using speech synthesis is generating an audio file with a selected ElevenLabs voice. The Skill creates these audio files specifically for integration into conversational agents.

How does the macOS say command UX compare to this ElevenLabs speech synthesis tool?

The macOS say command UX provides a simple command-line interface for immediate local playback, which this ElevenLabs speech synthesis tool mimics. It adds high-quality voice selection, pronunciation rules, and audio tag-based delivery modifications.

Are there limitations when applying pronunciation rules in text-to-speech generation?

Limitations in text-to-speech generation with pronunciation rules depend entirely on the ElevenLabs API capabilities. The Skill passes your defined delivery modifications and pronunciation rules to the API to generate the audio file.