alterlab-genai-audio-producer

Automate end-to-end AI audio production from raw recordings to finished deliverables.

9|1|Updated Mar 26, 2026
One-click install
npx skills add https://github.com/AlterLab-IEU/AlterLab-FC-Skills --skill alterlab-genai-audio-producer
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: alterlab-genai-audio-producer
Source: https://github.com/AlterLab-IEU/AlterLab-FC-Skills/tree/main/skills/genai/alterlab-genai-audio-producer
Command: npx skills add https://github.com/AlterLab-IEU/AlterLab-FC-Skills --skill alterlab-genai-audio-producer

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill provides an autonomous pipeline for transforming raw recordings into broadcast-ready audio, leveraging ElevenLabs tools to cleanup, generate narration, add SFX and music, and assemble the final mix.

Core Features & Use Cases

  • End-to-end audio production: cleanup with Voice Isolator, TTS narration, SFX, Eleven Music, and Studio 3.0 assembly.
  • Batch-ready workflows: support for content series with consistent voice, pacing, and loudness targets.
  • Export & transcription: export in multiple formats and generate Scribe v2 transcripts for show notes and accessibility.

Quick Start

Provide a complete end-to-end audio production plan for a scripted podcast episode using ElevenLabs tools.

Frequently Asked Questions about alterlab-genai-audio-producer

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate end-to-end AI audio production for a podcast?

You can automate end-to-end AI audio production by using a pipeline that applies Voice Isolator for cleanup, TTS for narration, and Studio 3.0 for final assembly. This handles everything from raw recordings to broadcast-ready podcast audio.

What loudness targets should I use for broadcast vs podcast audio?

For broadcast-ready audio, target -24 LUFS, while podcast production requires -16 LUFS. The workflow automatically satisfies these specific loudness targets alongside sample-rate and bit-depth consistency during the final assembly.

Can I generate transcripts from my TTS narration automatically?

Yes, Scribe v2 generates transcripts automatically during the audio production workflow. This provides accurate text outputs for show notes and accessibility alongside your exported MP3, WAV, or AAC audio files.

Does this AI audio workflow support adding sound effects and music?

The workflow supports adding sound effects and Eleven Music scoring natively. It integrates SFX generation and music scoring directly into the Studio 3.0 assembly process to produce a finished, mixed deliverable.

What audio formats can I export using this AI production pipeline?

You can export finished audio projects to MP3, WAV, and AAC formats. The pipeline ensures sample-rate and bit-depth consistency before exporting your final TTS narration and mixed audio deliverables.