acestep

Generate original music, vocal tracks, covers, and stem extractions from text prompts.

Updated Apr 2, 2026
One-click install
npx skills add https://github.com/gyalamanch001a/pur-new --skill acestep-gyalamanch001a
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: acestep
Source: https://github.com/gyalamanch001a/pur-new/tree/main/.agents/skills/acestep
Command: npx skills add https://github.com/gyalamanch001a/pur-new --skill acestep-gyalamanch001a

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill turns creative audio requests into usable music outputs without manual composition, helping teams produce background tracks, jingles, vocal songs, and stem edits quickly for content production.

Core Features & Use Cases

  • Original music generation from text prompts with control over mood, genre, instrumentation, BPM, key, and duration.
  • Vocal song creation with structured lyrics, plus cover and style-transfer workflows from reference audio.
  • Stem extraction for isolating vocals, drums, bass, guitar, piano, and other parts from mixed audio.
  • Video production use cases such as title music, problem-solution reveals, product demos, and call-to-action transitions.

Quick Start

Ask for an upbeat 60-second corporate background track with a clear mood, then have the Skill generate a polished audio file for your video.

Frequently Asked Questions about acestep

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate background music for video production from text prompts?

Generate background music for video production by providing structured text prompts that specify mood, genre, instrumentation, BPM, key, and duration. The Skill processes these inputs to produce polished 48 kHz MP3, WAV, or FLAC audio files suitable for title music, product demos, and call-to-action transitions.

Can I extract stems like vocals, drums, and bass from a mixed audio track?

Yes, you can extract stems to isolate vocals, drums, bass, guitar, piano, and other parts from mixed audio. This stem extraction capability allows you to separate individual instruments and vocal tracks for remixing, cover creation, or style-transfer workflows from reference audio.

What do I need to create vocal songs with structured lyrics and cover versions?

Creating vocal songs and cover versions requires structured lyrics formatting, reference audio for style transfer, and optional seed locking for consistency. You can control BPM and key settings to generate vocal tracks that match your desired musical style and produce high-quality 48 kHz audio outputs.

Does this music generation tool support scene-matched presets for soundtracks?

Yes, the music generation tool supports scene-matched presets for soundtracks, narration beds, and branded sonic assets. You can use these presets alongside prompt-driven audio requests to create scene-appropriate background music and jingles for various video production contexts.

How do I ensure consistent audio outputs across multiple generation sessions?

Ensure consistent audio outputs across multiple generation sessions by using optional seed locking alongside structured captions, lyrics formatting, and BPM and key controls. This approach maintains uniformity in mood, genre, and instrumentation when generating original music, vocal tracks, or cover versions.

What audio formats can I export when generating original music and stems?

When generating original music and stems, you can export audio in 48 kHz MP3, WAV, or FLAC formats. These high-quality output options support various content production needs, from background tracks and jingles to stem extractions and cover versions for video production.