sag

Convert text to speech using the ElevenLabs TTS engine with customizable voice options.

Updated Mar 10, 2026
One-click install
npx skills add https://github.com/lemonlqf/openclaw-rtsp --skill sag-lemonlqf
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sag
Source: https://github.com/lemonlqf/openclaw-rtsp/tree/main/skills/sag
Command: npx skills add https://github.com/lemonlqf/openclaw-rtsp --skill sag-lemonlqf

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill provides a convenient way to convert text into speech using ElevenLabs' advanced TTS engine, mimicking the familiar user experience of macOS's built-in 'say' command.

Core Features & Use Cases

  • Text-to-Speech Conversion: Generate natural-sounding speech from text using various ElevenLabs models.
  • Customizable Voice and Delivery: Control pronunciation, add pauses, and apply emotional tags for expressive speech.
  • Use Case: You need to create an audio narration for a presentation or a voice response for your AI assistant. Use this Skill to generate high-quality speech with specific emotional tones and pacing.

Quick Start

Use sag to speak the phrase "Hello there" with the default voice.

Frequently Asked Questions about sag

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert text to speech using ElevenLabs TTS?

You can convert text to speech by providing your text string and specifying customizable voice and delivery options. The Skill processes the input through the ElevenLabs engine to generate natural-sounding speech for local playback or audio file generation.

Do I need an ElevenLabs API key to generate voice audio?

Yes, an ElevenLabs API key is required for authentication and service access to generate voice audio. The Skill requires this key to interface with the ElevenLabs TTS engine and synthesize speech from your text.

Can I control pronunciation and add pauses for expressive speech synthesis?

Yes, you can control pronunciation, add pauses, and apply emotional tags for expressive speech synthesis. This allows you to generate audio narrations with specific emotional tones and pacing tailored to your use case.

Does this text-to-speech tool work like the macOS say command?

Yes, this text-to-speech tool mimics the familiar user experience of the macOS say command. It provides a mac-style say UX while utilizing the ElevenLabs advanced TTS engine to generate higher quality voice output.

What is the best way to create an audio narration for an AI assistant response?

The best way to create an audio narration is using a TTS engine with customizable delivery options. You can generate dynamic voice responses by synthesizing text into local audio files with specific emotional tones and pacing.

Can I save the generated speech as an audio file for local playback?

Yes, you can save the generated speech as an audio file for local playback. The Skill supports audio file generation alongside local playback, allowing you to store and use the synthesized voice output for applications requiring dynamic voice responses.