What problem does it solve?
This Skill automates the manual, repetitive, and fragile process of producing images, videos, edits, and speech from the Higgsfield web interface by driving Chrome and extracting results so users can get assets without using APIs or repeating UI steps.
Core Features & Use Cases
- Browser automation: Opens or switches to higgsfield.ai tabs, selects models, enters prompts, and clicks generate using the claude-in-chrome tools.
- Multi-modal generation: Supports text-to-image, text-to-video, image editing/inpaint, DoP image-to-video, Cinema Studio, speech generation, and specialized mini-app workflows with model selection.
- File upload and polling: Uploads local reference images, monitors generation progress with configurable timeouts, captures screenshots, and extracts download links or media URLs.
- Use Case: Produce a 4K photorealistic portrait with Seedream 4.5, animate a still image into a short clip, or generate expressive TTS audio using your logged-in Higgsfield account.
Quick Start
Use the higgsfield skill to generate a photorealistic portrait with seedream 4.5 by saying: /higgsfield seedream photorealistic portrait of a woman in golden hour