sag

Convert text to speech using the ElevenLabs API with a macOS say-style interface.

1|Updated Mar 8, 2026
One-click install
npx skills add https://github.com/syxscott/PaleoClaw --skill sag-syxscott
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sag
Source: https://github.com/syxscott/PaleoClaw/tree/main/skills/sag
Command: npx skills add https://github.com/syxscott/PaleoClaw --skill sag-syxscott

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a convenient way to convert text into speech using ElevenLabs' advanced TTS technology, mimicking the familiar user experience of macOS's 'say' command.

Core Features & Use Cases

  • Text-to-Speech Conversion: Generate natural-sounding speech from any text input.
  • Voice Customization: Select from various ElevenLabs voices and control pronunciation and delivery.
  • Use Case: Quickly generate audio responses for your AI agent, create voiceovers for presentations, or simply have text read aloud in a high-quality voice.

Quick Start

Use the sag skill to say "Hello there" with the default voice.

Frequently Asked Questions about sag

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert text to speech using the ElevenLabs API with a macOS say command interface?

Text-to-speech conversion is handled by sending text to the ElevenLabs API through a macOS 'say' command-like interface, generating natural-sounding speech. This allows you to quickly produce high-quality audio output from any text input.

Can I control voice pronunciation and delivery during voice synthesis?

Voice synthesis customization is supported through SSML or specific tags for enhanced vocal performance. You can select from various ElevenLabs voices and control pronunciation, normalization, and delivery for expressive speech generation.

Does this text-to-speech tool support multilingual and fast speech generation models?

Multilingual and fast speech generation are supported through various ElevenLabs models. The tool accommodates different use cases, allowing you to switch between models for expressive, multilingual, or rapid audio output based on your requirements.

What is the best way to generate voiceovers for presentations using ElevenLabs TTS?

Generating voiceovers for presentations is achieved by passing your script text into the ElevenLabs TTS interface. The system mimics the familiar macOS 'say' UX, providing a straightforward method to create high-quality voice audio for your slides.

Are there limitations when using SSML tags for text-to-speech normalization?

Text-to-speech normalization relies on ElevenLabs API capabilities and specific SSML tags for pronunciation control. While detailed vocal performance adjustments are supported, outputs are constrained by the selected ElevenLabs model's specific limitations and available voices.