elevenlabs

Convert text into spoken audio using ElevenLabs voices and models.

Updated Apr 16, 2026
One-click install
npx skills add https://github.com/yethikrishna/humble --skill elevenlabs-yethikrishna
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: elevenlabs
Source: https://github.com/yethikrishna/humble/tree/main/core/kortix-master/opencode/skills/GENERAL-KNOWLEDGE-WORKER/elevenlabs
Command: npx skills add https://github.com/yethikrishna/humble --skill elevenlabs-yethikrishna

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) components.

What problem does it solve?

Convert written content into high-quality spoken audio so you can automate narration, voiceovers, and ambient sound generation without recording humans or hiring voice talent.

Core Features & Use Cases

  • Text-to-Speech (TTS): Produce natural-sounding audio from plain text or files using multiple models and output formats.
  • Voice Cloning: Create custom voices from user-provided audio samples for personalized messages or branded narration.
  • Batch Processing & SFX: Convert long documents into multiple audio files, generate sound effects from prompts, and tune voice parameters like stability, similarity, style, and speed.
  • Use Case: Turn a product report into a narrated MP3 for a podcast intro, clone a user's voice for personalized notifications, or batch-convert chapters of an ebook into separate audio files.

Quick Start

Generate narration for workspace/report.md using the voice Rachel and save the output as report.mp3.

Frequently Asked Questions about elevenlabs

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I convert text documents into natural-sounding audio for batch audiobook generation?

Batch audiobook generation converts text documents into multiple natural-sounding audio files. This Skill handles file input and output, allowing you to batch-convert long documents or ebook chapters into separate spoken audio files with multiple voices and models.

Can I clone a custom voice from audio samples for podcast intros?

Yes, voice cloning creates custom voices from user-provided audio samples. You can clone a voice to generate personalized podcast intros, branded narration, or notifications without recording humans or hiring voice talent.

Do I need an ElevenLabs API key to generate text-to-speech and sound effects?

Yes, you need an ElevenLabs API key to generate text-to-speech audio and sound effects. The API key allows the Skill to access multiple models, handle file input and output, and expose parameters for stability, similarity, style, and speed.

What parameters can I tune for text-to-speech voice generation?

Text-to-speech voice generation exposes parameters for stability, similarity, style, speed, and output formats. These controls let you adjust the natural-sounding audio characteristics to suit narration, voiceovers, and ambient sound generation.

How do I synthesize sound effects from text prompts?

Sound effect synthesis generates ambient sounds from text prompts. The Skill processes your text prompts to produce short sound effects, handling the audio generation and output formatting directly without requiring manual recording.