What problem does it solve? Creating or editing video content normally requires filming, animation skills, or complex editing software. This Skill lets you generate physically realistic videos from text, images, or reference photos, and edit existing videos with natural language instructions, all through simple CLI commands. ## Core Features & Use Cases - Text-to-Video (T2V): Generate videos from text prompts at 720P or 1080P, up to 15 seconds, in multiple aspect ratios. - Image-to-Video and Reference-to-Video (I2V/R2V): Animate a still image or preserve characters from up to 9 reference images for consistent multi-character scenes. - Natural Language Video Editing: Modify existing videos (backgrounds, characters, weather, audio) using plain text instructions with optional reference images. - Use Case: A marketing team needs a 10-second product demo clip with a consistent brand character. They provide a character reference photo and a scene prompt to the R2V model, then refine the result with the Video Edit model. ## Quick Start Ask the AI to generate a 10-second 1080P video of a golden retriever running through autumn leaves using the HappyHorse text-to-video model.