sag

Generate spoken audio from text using ElevenLabs TTS models.

10|Updated Mar 12, 2026
One-click install
npx skills add https://github.com/wixette/clawnotes --skill sag-wixette
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sag
Source: https://github.com/wixette/clawnotes/tree/main/openclaw-snapshots/20260312/skills/sag
Command: npx skills add https://github.com/wixette/clawnotes --skill sag-wixette

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a convenient way to generate spoken audio from text using ElevenLabs' advanced text-to-speech technology, mimicking the user experience of the macOS say command.

Core Features & Use Cases

  • Text-to-Speech Generation: Convert written text into natural-sounding speech using various ElevenLabs models.
  • Voice Customization: Select specific voices, control pronunciation, and adjust delivery with tags for tone and emotion.
  • Use Case: Generate an audio response for a user request, such as explaining a complex topic in a specific voice character like a "crazy scientist."

Quick Start

Use the sag skill to speak the phrase "Hello there" with the default voice.

Frequently Asked Questions about sag

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate spoken audio from text using ElevenLabs TTS?

To generate spoken audio from text using ElevenLabs TTS, this tool provides a macOS-style command-line interface that converts written text into natural-sounding speech. It supports various models and outputs audio files for chat media embedding.

Can I select specific voices and control pronunciation during speech synthesis?

Yes, during speech synthesis you can select specific voices and adjust delivery. The tool supports advanced pronunciation and delivery controls via SSML or custom tags to define tone and emotion for the generated audio.

Does the sag skill work like the macOS say command for text-to-speech?

Yes, the sag skill works like the macOS say command for text-to-speech generation. It mimics that user experience while utilizing ElevenLabs' advanced speech synthesis technology to produce the spoken audio files.

What is the best way to add voice responses to chat applications using ElevenLabs?

The best way to add voice responses to chat applications using ElevenLabs is to generate audio files from text for media embedding. You can use custom tags to voice a specific character like a crazy scientist.

Are there limitations when using SSML or custom tags for voice generation?

When using SSML or custom tags for voice generation, limitations depend on the supported ElevenLabs models. You must ensure your chosen model supports the specific pronunciation and delivery controls you apply to the text.