clawvox

Generate speech, transcribe audio, clone voices, and dub audio via the ElevenLabs API.

2|Updated Jan 22, 2026
One-click install
npx skills add https://github.com/aztr0nutzs/NET_NiNjA.v1.2 --skill clawvox
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: clawvox
Source: https://github.com/aztr0nutzs/NET_NiNjA.v1.2/tree/main/skills/skills-folders/clawvox
Command: npx skills add https://github.com/aztr0nutzs/NET_NiNjA.v1.2 --skill clawvox

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires curl, jq, and includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill transforms your AI assistant into a powerful voice production platform, enabling lifelike speech generation, accurate transcription, voice cloning, and advanced audio manipulation.

Core Features & Use Cases

  • Text-to-Speech (TTS): Generate natural-sounding speech from text using various voices and models.
  • Speech-to-Text (STT): Transcribe audio files with high accuracy, supporting multiple languages.
  • Voice Cloning: Create custom voices from audio samples.
  • Sound Effects & Dubbing: Generate sound effects from descriptions and translate audio into multiple languages.
  • Use Case: A content creator can use ClawVox to generate voiceovers for videos, transcribe interviews, and even clone their own voice for consistent audio branding across projects.

Quick Start

Use ClawVox to speak the phrase "Hello from ElevenLabs Voice Studio" using the Adam voice.

Frequently Asked Questions about clawvox

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate natural-sounding text-to-speech audio from written content?

Text-to-speech generation uses the ElevenLabs API to convert written text into natural-sounding audio. You can customize the output by selecting specific voices, models, and parameters to match your desired tone and style.

Can I clone a voice from an existing audio sample?

Yes, voice cloning creates custom voices from provided audio samples. This allows you to generate consistent speech and audio branding using a specific person's voice characteristics.

How do I transcribe speech-to-text from an audio file?

Speech-to-text transcription processes audio files to convert spoken words into written text. This functionality supports transcribing audio with high accuracy across multiple languages.

Do I need an ElevenLabs API key to use voice dubbing and sound effects?

Yes, an ElevenLabs API key is required to use multilingual audio dubbing and sound effect generation. You also need command-line tools like curl and jq installed to execute the operations.

What is the best way to translate and dub audio into multiple languages?

Multilingual audio dubbing translates spoken audio into different languages by processing the original file through the ElevenLabs API. This generates localized audio outputs for diverse audiences.