What problem does it solve?
ElevenLabs audio mastery skill with narration generation, timestamp extraction, WebVTT companion output, manifest-driven narration write-back, pronunciation dictionary creation, dialogue generation, sound effects, and music composition through the shared ElevenLabsClient. This skill serves as the voice-production layer in the repo's three-layer architecture: Voice Director (agent judgment) -> elevenlabs-audio (skill - tool expertise) -> ElevenLabsClient (API client - connectivity).
Core Features & Use Cases
- Narration generation with optional timestamps and corresponding WebVTT output.
- Manifest-driven narration workflows with write-back to lesson assets.
- Pronunciation dictionary creation and management for medical terminology.
- Dialogue generation, sound effects (SFX), and music composition wrappers.
- Style-guide-driven defaults and voice-preview integration to support consistent production.
Quick Start
Generate a narrated lesson using ElevenLabs with timestamps, pronunciation dictionaries, dialogue, SFX, and music wrappers.