elevenlabs

Convert text into speech and generate sound effects via the ElevenLabs API.

Updated Apr 8, 2026
One-click install
npx skills add https://github.com/aimentor606/aether --skill elevenlabs-aimentor606
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: elevenlabs
Source: https://github.com/aimentor606/aether/tree/main/core/kortix-master/opencode/skills/GENERAL-KNOWLEDGE-WORKER/elevenlabs
Command: npx skills add https://github.com/aimentor606/aether --skill elevenlabs-aimentor606

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

Convert written text and prompts into high-quality spoken audio and sound effects so teams can produce narration, voiceovers, and personalized audio without studio recording or manual editing.

Core Features & Use Cases

  • Text-to-Speech: Generate natural, multilingual speech from text with selectable models and output formats.
  • Voice Cloning: Create custom voices from audio samples and reuse them for personalized messages or branded narration.
  • Batch Processing & SFX: Convert entire documents into audio files, split by paragraphs, and create ambient or effect audio from prompts.
  • Use Case: Turn a product guide into an audiobook, create podcast intros in a branded voice, or generate accessibility narration for documents and slides.

Quick Start

Generate a narrated MP3 of the file report.md using voice Rachel and save it as report_narration.mp3.

Frequently Asked Questions about elevenlabs

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert a text document into natural-sounding speech for narration?

To convert text into natural-sounding speech, this tool processes written documents and outputs high-quality spoken audio files. It supports selectable models and voice tuning parameters to produce narration without manual editing.

Can I generate audio from multiple text files at once for batch processing?

Yes, batch processing is supported for audio generation. You can convert entire documents into audio files, automatically splitting the text by paragraphs to produce multiple standard audio segments.

Do I need an ElevenLabs API key to generate voiceovers and sound effects?

Yes, an ElevenLabs API key is required to authenticate requests for generating voiceovers and sound effects. You must provide this key to access the text-to-speech, voice cloning, and audio generation features.

How does voice cloning work for creating personalized audio?

Voice cloning works by creating custom voices from provided audio samples. Once cloned, these voices can be reused for personalized messages or branded narration across your text-to-speech outputs.

What is the best way to generate ambient sound effects from a text prompt?

The best way to generate sound effects from a prompt is using the built-in SFX feature. It converts descriptive text prompts into ambient or effect audio, supplementing your standard text-to-speech generation.

Can I use text-to-speech to create accessibility narration for slides and documents?

Yes, you can use text-to-speech to create accessibility narration for slides and documents. It converts written content into high-quality spoken audio, eliminating the need for studio recording or manual editing.