What problem does it solve?
Choosing and configuring the right provider for MiniMax H3 (Hailuo 3.0) video generation is confusing because each platform uses different model identifiers, durations, and resolutions. This Skill routes text-to-video, image-to-video, first/last-frame, and reference-conditioned video requests to the correct provider with the correct parameters.
Core Features & Use Cases
- Multi-provider routing: Generate video through the MiniMax v2 API, fal.ai, Runway, ComfyUI Partner Nodes, or local open weights in ComfyUI, each with its correct model identifier.
- Multiple generation modes: Supports text-to-video, image-to-video from a first frame, first/last-frame animation, and reference-to-video combining images, video, and audio.
- Local offline generation: Run the open-weight Hailuo 3.0 stack locally in ComfyUI with the official workflow, diffusion model, Qwen3-VL text encoder, and VAEs.
- Use Case: Provide a product photo as a first frame and a prompt describing camera motion, and receive a 4-15 second 2K clip with native audio via the MiniMax direct API.
Quick Start
Ask the agent to generate a 10-second MiniMax H3 video from a text prompt describing the subject, camera path, and lighting using the MiniMax direct route.