audiocraft-audio-generation

Generate audio from text prompts using AudioCraft's MusicGen, AudioGen, and EnCodec pipelines.

Updated Jun 11, 2026
One-click install
npx skills add https://github.com/LamseyahElias/jarvis-cloud-v2 --skill audiocraft-audio-generation-lamseyahelias
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: audiocraft-audio-generation
Source: https://github.com/LamseyahElias/jarvis-cloud-v2/tree/main/hermes-agent/skills/mlops/models/audiocraft
Command: npx skills add https://github.com/LamseyahElias/jarvis-cloud-v2 --skill audiocraft-audio-generation-lamseyahelias

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

Audio content creation from natural language prompts is automated, enabling rapid music and sound design without manual composition.

Core Features & Use Cases

  • Generate music or sound effects from text prompts using MusicGen, AudioGen, and EnCodec.
  • Melody-conditioned or stereo outputs for immersive audio production.
  • Use cases include game audio, film sound design, podcasts, and rapid concept prototyping.

Quick Start

Describe your desired audio prompt and run the model to generate a short audio clip.

Frequently Asked Questions about audiocraft-audio-generation

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate music and sound effects from text prompts?

Generate music and sound effects from text prompts by applying AudioCraft's MusicGen, AudioGen, and EnCodec pipelines. Describe your desired audio prompt and run the model to produce a short audio clip for your multimedia project.

Can I use melody conditioning for audio generation in game sound design?

Melody conditioning for audio generation in game sound design is supported through MusicGen pipelines. You can produce melody-conditioned or stereo outputs to create immersive audio tailored to your specific game environment.

What do I need to configure to generate audio clips with AudioCraft pipelines?

To generate audio clips with AudioCraft pipelines, you need to configure appropriate model variants, sample rates, durations, and CFG parameters. These settings ensure the production of usable audio outputs for your projects.

Does text to audio generation work for rapid concept prototyping in film?

Text to audio generation works effectively for rapid concept prototyping in film. Automating audio content creation from natural language prompts enables fast sound design without requiring manual composition.

What is the best way to automate audio content creation from natural language?

The best way to automate audio content creation from natural language is using AudioCraft's MusicGen and AudioGen pipelines. This approach handles model variants and CFG parameters to produce usable audio outputs.

What are the limitations of text to audio generation for multimedia projects?

Limitations of text to audio generation for multimedia projects include the need to select appropriate model variants and manually configure sample rates, durations, and CFG parameters to yield usable audio outputs.