audiocraft

Generate music and sound effects from text using Meta's AudioCraft library.

Updated Jul 3, 2026
One-click install
npx skills add https://github.com/Toqsick/MaxClaw --skill audiocraft-toqsick
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: audiocraft
Source: https://github.com/Toqsick/MaxClaw/tree/main/.claude/skills/audiocraft
Command: npx skills add https://github.com/Toqsick/MaxClaw --skill audiocraft-toqsick

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires audiocraft, torch, transformers, and includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

This Skill enables users to generate music and sound effects from text descriptions, streamlining the creation of audio content for various applications.

Core Features & Use Cases

  • Text-to-Music: Convert text descriptions into music using MusicGen, with features like melody conditioning and style transfer.
  • Text-to-Sound: Generate sound effects and environmental audio from text descriptions using AudioGen.
  • Use Case: Create a personalized audio track for a presentation by inputting a description like "upbeat electronic dance music with a strong beat."

Quick Start

Use the audiocraft skill to generate music from the text description 'upbeat electronic dance music with a strong beat'.

Frequently Asked Questions about audiocraft

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate music from a text description?

To generate music from text, use the audiocraft skill to process descriptions like 'upbeat electronic dance music with a strong beat' and produce audio tracks through MusicGen, enabling melody conditioning and style transfer for content creation.

Can I create sound effects from text using AudioGen?

Yes, you can generate sound effects and environmental audio from text descriptions using AudioGen, streamlining audio production for interactive applications and content creation.

Do I need Python libraries like torch and transformers to generate text-to-music?

Yes, generating text-to-music and text-to-sound requires Python libraries including audiocraft, torch, and transformers to run Meta's AudioCraft library for your audio production workflow.

What is the best way to convert text to music for a presentation?

The best way to convert text to music for a presentation is to input a specific text description into the audiocraft skill, which uses MusicGen to output a personalized audio track.

Does MusicGen support melody conditioning and style transfer?

Yes, MusicGen supports melody conditioning and style transfer features, allowing you to convert text descriptions into customized music tracks for content creation.

What are the limitations of using text-to-sound generation for audio production?

While text-to-sound generation streamlines audio content creation, its limitations depend on the available Python environment and the specificity of your text descriptions when using AudioGen and MusicGen models.