What problem does it solve?
This Skill provides a comprehensive interface to the ElevenLabs API, enabling users to generate high-quality audio from text, manipulate existing audio, and transcribe speech, all through a simple command-line wrapper.
Core Features & Use Cases
- Text-to-Speech (TTS): Convert text into natural-sounding speech with various voices and languages.
- Voice Cloning: Create custom voices from audio samples.
- Audio Generation: Produce sound effects and music from text descriptions.
- Audio Manipulation: Perform speech-to-speech conversion, isolate vocals, and dub audio/video.
- Transcription: Convert spoken audio into text.
- Use Case: A content creator can use this Skill to generate voiceovers for videos, create character dialogue for a game, or even clone their own voice for consistent narration across multiple projects.
Quick Start
Use the elevenlabs skill to convert the text 'Hello, world!' to speech using the 'Rachel' voice and save it to 'hello.mp3'.