Podcast Generate

Convert user content or web search results into podcast scripts and WAV audio.

Updated Apr 28, 2026
One-click install
npx skills add https://github.com/ncsound919/deterministic-brain --skill podcast-generate-ncsound919
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: Podcast Generate
Source: https://github.com/ncsound919/deterministic-brain/tree/main/skills/podcast-generate
Command: npx skills add https://github.com/ncsound919/deterministic-brain --skill podcast-generate-ncsound919

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill automates turning input materials or web-sourced information into a ready-to-publish podcast script and accompanying audio, saving time and ensuring consistency.

Core Features & Use Cases

  • Converts documents or web search results into a dual-host or single-host podcast script with structured dialogue and natural pacing.
  • Generates a final podcast screenplay (podcast_script.md) and audio file (podcast.wav) using z-ai web SDK for both scripting and TTS.
  • Useful for content creators, educators, and teams needing reproducible audio content from text or research.

Quick Start

Run the generate command with your input file or topic to produce podcast_script.md and podcast.wav.

Frequently Asked Questions about Podcast Generate

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate a podcast script from text content automatically?

To generate a podcast script from text, the Skill converts your input documents or web-search results into a structured dialogue. It outputs a ready-to-publish podcast_script.md file using natural pacing and dual-host or single-host formats.

Can I convert web search results directly into podcast audio?

Yes, you can convert web search results directly into podcast audio. The Skill processes web-sourced information to generate both a structured podcast_script.md and a final podcast.wav audio file using TTS.

How does the text-to-speech system handle podcast duration limits?

The text-to-speech system auto-adjusts the podcast duration from 3 to 20 minutes based on content length. It enforces output structure, speaker alternation, and validation rules during audio generation.

Do I need to format my input materials before generating a podcast?

No extensive formatting is required before generating a podcast. You can provide raw documents or topics, and the Skill will automatically structure the content into a dual-host or single-host dialogue with natural pacing.

What is the difference between dual-host and single-host podcast generation?

Dual-host podcast generation structures content as an alternating dialogue between two speakers, while single-host uses a monologue format. The Skill enforces speaker alternation rules for dual-host and outputs a validated script.

Are there limitations when converting long research documents into a podcast?

When converting long research documents into a podcast, the main limitation is the auto-adjusted duration cap of 20 minutes. Content exceeding this length will be constrained by the validation rules enforced during script generation.