What problem does it solve?
Quickly produce high-quality audio narrations of blog posts so readers can listen instead of read, enabling accessibility, podcast repurposing, and on-site audio embeds without manual TTS engineering.
Core Features & Use Cases
- Multi-mode narration: summary (200-300 words), full article read-aloud, or two-speaker dialogue for short podcast-style episodes.
- Voice catalog & pairing: 30 prebuilt Gemini voices with recommended pairings for host/expert dialogue and content-type suggestions.
- Operational outputs: MP3 (or WAV fallback), HTML5 embed code, duration & cost estimates, and placement guidance for blog platforms.
- Robust workflow: local venv management, dry-run cost estimates, FFmpeg fallback, and graceful silent return when API key is missing so writing workflows are never blocked.
Quick Start
Ask the skill to generate audio by saying: /blog audio generate my-article.md --mode summary --voice Charon