What problem does it solve?
This Skill helps media teams separate existing mixed recordings into usable audio stems without confusing source separation with audio generation, mixing, or mastering. It provides the operational judgment needed to select appropriate targets, manage costs and API limits, validate output quality, and handle rights, consent, and privacy concerns.
Core Features & Use Cases
- Music Stem Separation: Isolate vocals, drums, bass, guitar, piano, keys, strings, winds, instrumental beds, and residual components for karaoke, remixing, practice tracks, and immersive re-authoring.
- Post-Production and Speech Workflows: Create dialogue, music-and-effects, effects, and per-speaker stems for localization, dubbing, interviews, podcasts, and archival restoration.
- Production Validation: Review stems for bleed, watery artifacts, transient smearing, missing energy, and phase problems, then apply appropriate repair or reprocessing strategies.
- API and Platform Guidance: Operate the Tasks API, choose between polling and webhooks, estimate credit usage, manage expiring downloads, and decide when the on-device SDK or a local model such as Demucs is more suitable.
- Use Case: Prepare a licensed episode for dubbing by extracting a clean dialogue reference and a music-and-effects bed, then verify that the resulting stems reconstruct the original mix without significant bleed or missing energy.
Quick Start
Use the AudioShake stem separation skill to create the dialogue and music-and-effects stems for my licensed episode, estimate the credits, explain the retrieval workflow, and provide a broadcast-readiness checklist.