What problem does it solve?
This Skill helps media teams create production-ready music, sound effects, ambience, foley, loops, and sonic-branding assets with Stable Audio while avoiding unsuitable speech workflows, rights risks, and unreliable delivery decisions.
Core Features & Use Cases
- Model and API Routing: Choose between hosted Stable Audio 2, 2.5, and 3 endpoints or open-weight models based on duration, latency, quality, and local-data requirements.
- Audio Generation and Editing: Plan text-to-audio, audio-to-audio, inpainting, and continuation workflows with appropriate prompts, durations, steps, seeds, formats, and source-audio strength.
- Production Guardrails: Enforce rights-cleared uploads, document provenance, handle asynchronous Stable Audio 3 polling, and review audio for artifacts, editability, loudness, continuity, and delivery readiness.
- Use Case: Create a 40-second optimistic instrumental product-launch bed under narration, or generate multiple variants of a rights-cleared game chime while preserving its core motif.
Quick Start
Use the stable-audio skill to create a rights-safe 30-second SaaS launch music bed with a narration-friendly arrangement, generation parameters, provenance notes, and a final audio QA checklist.