podcast-generator

Generate conversational podcast scripts and audio from research documents.

Updated Apr 28, 2026
One-click install
npx skills add https://github.com/Kirankumar2604/solutionChallenge --skill podcast-generator-kirankumar2604
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: podcast-generator
Source: https://github.com/Kirankumar2604/solutionChallenge/tree/main/.local/secondary_skills/podcast-generator
Command: npx skills add https://github.com/Kirankumar2604/solutionChallenge --skill podcast-generator-kirankumar2604

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires elevenlabs, pydub, ffmpeg-normalize.

What problem does it solve?

This skill solves the challenge of converting dense research, articles, or complex topics into engaging, conversational audio content without requiring professional studio equipment or manual scriptwriting.

Core Features & Use Cases

  • Conversational Scripting: Generates natural, host-guest dialogue scripts using a proven two-host methodology.
  • Audio Production: Integrates with ElevenLabs to transform scripts into high-quality, normalized MP3 audio files.
  • Series Management: Supports long-term content planning with show bibles, episode calendars, and recurring segment structures.
  • Use Case: A researcher can input a complex technical paper and generate a 15-minute conversational podcast episode that explains the findings to a general audience.

Quick Start

Use the podcast-generator skill to create a conversational script and audio episode about the latest advancements in quantum computing based on the provided research notes.

Frequently Asked Questions about podcast-generator

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I turn research notes into a conversational podcast?

To convert research notes into a conversational podcast, provide source documents to generate host-guest dialogue scripts in solo, duo, interview, or narrative formats. The tool then uses ElevenLabs text-to-speech integration and loudness normalization to output production-ready audio.

Do I need an ElevenLabs API key to generate text-to-speech audio?

Yes, an ElevenLabs API key is required to generate text-to-speech audio. The process relies on ElevenLabs to transform conversational scripts into high-quality MP3 files, followed by automated loudness normalization using pydub and ffmpeg-normalize.

Can I create a multi-episode podcast series from technical papers?

Yes, you can create a multi-episode podcast series from technical papers using series management features. This supports long-term content planning by generating show bibles, episode calendars, and recurring segment structures for ongoing audio content.

What podcast script formats are supported for audio content generation?

Supported podcast script formats for audio content generation include solo, duo, interview, and narrative styles. These formats use a proven two-host methodology to create natural conversational scripts from dense research material or complex articles.

Does the audio production pipeline handle loudness normalization for MP3 files?

Yes, the audio production pipeline handles loudness normalization for MP3 files. After generating text-to-speech audio via ElevenLabs, the process applies ffmpeg-normalize and pydub libraries to ensure the final podcast output meets production-ready loudness standards.

What is the best way to explain complex research findings to a general audience?

The best way to explain complex research findings to a general audience is generating a conversational podcast episode. Transforming dense technical papers into natural host-guest dialogue scripts produces engaging audio content without requiring manual scriptwriting or professional studio equipment.