ai-podcast-creation

Generate podcasts by synthesizing text-to-speech audio, composing music, and merging audio elements.

688|95|Updated Jan 31, 2026
One-click install
npx skills add https://github.com/inference-sh/skills --skill ai-podcast-creation-inference-sh
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: ai-podcast-creation
Source: https://github.com/inference-sh/skills/tree/main/guides/content/ai-podcast-creation
Command: npx skills add https://github.com/inference-sh/skills --skill ai-podcast-creation-inference-sh

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes scripts (resource) and references (resource) components.

What problem does it solve?

This Skill automates the creation of AI-powered podcasts, simplifying the process of generating audio content with text-to-speech, music, and editing.

Core Features & Use Cases

  • Text-to-Speech: Generate speech from text using various AI voices (Kokoro TTS, DIA TTS, Chatterbox).
  • AI Music Generation: Create custom intro/outro music and background tracks.
  • Audio Merging & Editing: Combine voice segments, music, and sound effects into a cohesive podcast episode.
  • Use Case: Produce professional-sounding podcasts, audiobooks, or audio newsletters efficiently, even with multiple AI-generated voices for conversations.

Quick Start

Use the ai-podcast-creation skill to generate a podcast segment with the prompt "Welcome to the AI Frontiers podcast. Today we explore the latest developments in generative AI." using the am_michael voice.

Frequently Asked Questions about ai-podcast-creation

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate a podcast using text-to-speech and AI voices?

AI podcast creation generates audio by synthesizing text-to-speech with voices like Kokoro TTS, DIA TTS, or Chatterbox. You provide the text script, and the system produces voice segments for your podcast automatically.

Can I create multi-voice conversations for an AI-generated podcast?

Yes, AI podcast creation supports multi-voice conversations. You can assign different AI voices to various text segments, allowing automated dialogue generation for interviews or panel-style podcast episodes.

How does AI music generation work for podcast intro and outro tracks?

AI music generation creates custom intro, outro, and background tracks for podcasts. The system composes original audio segments that are then merged with voice segments to produce a complete episode.

What is the best way to merge voice segments and background music into a full podcast episode?

The best way to merge audio elements is using automated audio merging and editing. The system combines generated voice segments, AI music, and sound effects into a cohesive podcast episode ready for distribution.

Can I use this automated podcast creation workflow for audiobook narration?

Yes, automated podcast creation supports audiobook narration and audio newsletter generation. The text-to-speech and audio merging workflows apply directly to producing long-form spoken audio from text scripts.

Do I need any external audio editing software to produce an AI podcast?

No, you do not need external audio editing software. The AI podcast creation Skill handles text-to-speech synthesis, AI music generation, and audio merging internally to output a finished podcast episode.