sag

Synthesize ElevenLabs text-to-speech audio with local playback.

Updated Mar 26, 2026
One-click install
npx skills add https://github.com/tedtv1007-ctrl/milk-skills-library --skill sag-tedtv1007-ctrl
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: sag
Source: https://github.com/tedtv1007-ctrl/milk-skills-library/tree/main/sag
Command: npx skills add https://github.com/tedtv1007-ctrl/milk-skills-library --skill sag-tedtv1007-ctrl

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires sag.

What problem does it solve?

This Skill bridges the gap between text-based AI responses and high-quality, expressive audio output by providing a seamless interface to ElevenLabs text-to-speech services.

Core Features & Use Cases

  • Expressive TTS: Supports advanced voice tags like whispers, laughs, and dramatic pauses for natural-sounding speech.
  • Flexible Playback: Enables direct local audio generation and playback from the command line.
  • Use Case: Use this to generate a voice-over for a presentation or to have the AI read complex technical documentation aloud in a specific character voice.

Quick Start

Use the sag skill to generate an audio file named greeting.mp3 using the Clawd voice with the text Hello there.

Frequently Asked Questions about sag

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate expressive text-to-speech audio with whispers and pauses from the command line?

To generate expressive text-to-speech audio with whispers and pauses from the command line, use this Skill to interface with ElevenLabs TTS. It supports advanced voice tags for natural-sounding speech and enables local audio playback directly through the sag CLI binary.

Do I need an ElevenLabs API key to use text-to-speech voice synthesis locally?

Yes, you need a valid ElevenLabs API key to use text-to-speech voice synthesis locally. The Skill requires this key alongside the sag CLI binary to execute speech synthesis and manage voice-driven audio generation tasks on your system.

Can I use a specific character voice to read technical documentation aloud?

You can use a specific character voice to read technical documentation aloud by leveraging this Skill's ElevenLabs integration. It facilitates character-driven narration and automated audio content generation, outputting speech in your chosen voice profile.

How do I play ElevenLabs TTS audio locally after generating speech from text?

To play ElevenLabs TTS audio locally after generating speech from text, execute the sag CLI binary provided by this Skill. It facilitates direct local audio generation and playback, outputting synthesized voice files like greeting.mp3 to your system.

What is the best way to automate voice-based AI interactions using ElevenLabs speech synthesis?

The best way to automate voice-based AI interactions using ElevenLabs speech synthesis is through this Skill's command line interface. It bridges text-based AI responses and high-quality audio output, enabling automated voice generation via the sag binary.

Does the sag CLI binary support automated audio content generation for presentations?

Yes, the sag CLI binary supports automated audio content generation for presentations. This Skill enables you to generate voice-overs by synthesizing text into expressive speech, providing high-quality audio output files for your presentation slideshows.