higgsfield-audio

Guides audio prompting for dialogue, lip-sync, SFX, ambient sound, and BGM in Higgsfield video generation.

Updated Jul 15, 2026
One-click install
npx skills add https://github.com/executiveusa/buffer-blaster- --skill higgsfield-audio-executiveusa
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: higgsfield-audio
Source: https://github.com/executiveusa/buffer-blaster-/tree/main/skills/higgsfield/skills/higgsfield-audio
Command: npx skills add https://github.com/executiveusa/buffer-blaster- --skill higgsfield-audio-executiveusa

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve? AI video models handle audio inconsistently — lip-sync drifts, uploaded tracks get overridden by generative audio, and prompts without audio direction produce silent or mismatched sound. This Skill provides model-specific rules for writing audio prompts that produce synchronized dialogue, sound effects, ambience, and music in Higgsfield-generated video. ## Core Features & Use Cases - Four-layer audio prompting: Structures dialogue, SFX, ambient, and BGM cues with per-model syntax for Kling 3.0, Seedance 1.5 Pro/2.0, Veo 3/3.1, and Grok Imagine Video. - Lip-sync discipline: Enforces 3–8 second dialogue clips, single-face framing, locked cameras, and per-language word budgets to prevent desync. - Audio-as-conditioning: Explains Seedance 2.0 @Audio1 reference usage for beat sync, timestamped [AUDIO: Xs] script blocks, and scoped melody/voice references. - Use Case: A creator generating a 15-second Seedance 2.0 clip uploads a trimmed MP3 hook as @Audio1, anchors it with a timestamp phrase, and writes a [AUDIO: Xs] dialogue block so the character's lines lip-sync while cuts land on the beat. ## Quick Start Ask the assistant to write a Higgsfield video prompt with dialogue, ambient sound, and beat-synced music for a specific model such as Seedance 2.0 or Kling 3.0.

Frequently Asked Questions about higgsfield-audio

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I add dialogue and lip-sync to AI-generated video?

Put quoted dialogue in the prompt with speaker and tone cues, keep clips 3–8 seconds, use medium close-up framing with one speaking face, and lock the camera. Remove head-motion tokens like nodding or turning, which compete with the lip-sync engine and cause desync.

Which AI video models support native audio generation?

Kling 3.0, Seedance 2.0, Seedance 1.5 Pro, Veo 3/3.1, and Grok Imagine Video generate audio jointly with video in one pass. Other models require adding audio in post with tools like Lipsync Studio.

How do I use an uploaded MP3 as a beat-sync reference in Seedance 2.0?

Attach the MP3 as @Audio1 and write an explicit audio-to-visual mapping, such as syncing camera cuts to downbeats and movement to the dynamic contour. Extract a 15-second build-to-drop segment at 256kbps or higher, since the model only reads the first 15 seconds.

Why does the model replace my uploaded audio track?

Ambient, SFX, or music tokens in the prompt invite the generative audio engine to override the uploaded file. Add a timestamp anchoring phrase like "Audio @Audio1 plays exactly as uploaded from 0s to end" and remove all other audio tokens.

What audio formats does Seedance 2.0 accept?

Seedance 2.0 accepts MP3 only — WAV, AAC, OGG, and FLAC fail silently with no error. Limits are 15 seconds per clip, 3 audio files, 10MB each, with 256kbps or higher recommended for beat detection.

When should I skip audio direction in a video prompt?

Skip audio cues when using models without native audio, when content is purely visual like product shots or landscapes, when audio will be added entirely in post, or when the prompt is at its word cap and visual direction matters more.