happyhorse-1-0

Generate text-to-video content via the RunComfy CLI using the HappyHorse 1.0 endpoint.

12|2|Updated May 18, 2026
One-click install
npx skills add https://github.com/runcomfy-com/skills --skill happyhorse-1-0
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: happyhorse-1-0
Source: https://github.com/runcomfy-com/skills/tree/main/happyhorse-1-0
Command: npx skills add https://github.com/runcomfy-com/skills --skill happyhorse-1-0

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Generate high-quality text-to-video shots without manually dealing with model selection, routing, or endpoint details, while keeping audio and character consistency aligned to the same generation pass.

Core Features & Use Cases

  • Text-to-video generation on RunComfy: Produces videos via the HappyHorse 1.0 text-to-video endpoint with a prompt-driven workflow.
  • Native 1080p and in-pass synchronized audio: Creates 1080p output and keeps audio (dialogue/ambient/Foley) synchronized within the same generation.
  • Multi-shot character consistency guidance: Helps structure prompts for multi-shot continuity using anchor-based descriptions.

Use case: Create a short multilingual ad or a talking-head style explainer where the voiceover and ambient audio match the described scene flow.

Quick Start

Ask for a short, camera-directed, multi-shot video prompt and specify the desired aspect ratio and duration, then generate via the local RunComfy CLI for the HappyHorse 1.0 text-to-video endpoint.

Frequently Asked Questions about happyhorse-1-0

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate 1080p text-to-video with synchronized audio?

To generate 1080p text-to-video with synchronized audio, provide a camera-directed, multi-shot prompt specifying aspect ratio and duration. The model routes requests to the HappyHorse 1.0 endpoint, producing video and matching ambient or dialogue audio in a single pass.

What is the best way to maintain character consistency across multiple video shots?

To maintain character consistency across multiple video shots, use anchor-based descriptions within your text-to-video prompt. This guides the model to keep character features aligned throughout storyboard-driven, camera-directed animation scenarios.

Can I create multilingual video content with matching voiceover and ambient audio?

Yes, you can create multilingual video content with matching voiceover and ambient audio. The system generates synchronized audio during the same generation pass, ensuring dialogue and ambient sounds match the described scene flow.

Do I need the RunComfy CLI to run the text-to-video endpoint?

Yes, you need the RunComfy CLI installed locally to run the text-to-video endpoint. The system invokes the CLI with a JSON prompt payload and returns downloaded result URLs from the CLI output directory.

What video settings can I configure for 1080p text-to-video workflows?

For 1080p text-to-video workflows, you can configure aspect ratio, resolution, duration, seed, and watermark settings. These parameters are passed in the JSON prompt payload to control the final video output format.

Why use a storyboard-driven approach for text-to-video generation?

A storyboard-driven approach for text-to-video generation helps structure prompts for multi-shot continuity. By using camera-directed anchor descriptions, you achieve better scene flow and visual consistency across generated 1080p video clips.