video-gen

Generate videos from images and text prompts using Higgsfield DOP and fal.ai Kling 3.0.

267|63|Updated Feb 15, 2026
One-click install
npx skills add https://github.com/modu-ai/cowork-plugins --skill video-gen-modu-ai
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: video-gen
Source: https://github.com/modu-ai/cowork-plugins/tree/main/moai-media/skills/video-gen
Command: npx skills add https://github.com/modu-ai/cowork-plugins --skill video-gen-modu-ai

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This Skill simplifies the production of high-quality videos by combining image and text inputs with advanced AI models, reducing the need for complex video editing skills.

Core Features & Use Cases

  • Image-to-Video Conversion: Transform static images into cinematic motion videos with customizable motion presets.
  • Text-to-Video Generation: Generate dynamic videos from textual prompts, suitable for marketing or storytelling.
  • Lip-Sync and Character Scenes: Create videos with animated characters speaking or lip-syncing based on images and prompts.
  • Use Case: For marketing, a user can turn a product image into an engaging promotional video using preset camera movements like orbit or dolly.

Quick Start

Provide an image URL and prompt describing the desired video motion to generate a cinematic video from your image.

Frequently Asked Questions about video-gen

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
Can I create text-to-video animations for marketing content?

You can generate a dynamic video from a text prompt using fal.ai Kling 3.0. This text-to-video generation feature creates dynamic sequences suitable for marketing campaigns or visual storytelling.

Does lip-sync video generation work with fal.ai Kling 3.0?

Creating dynamic motion videos from images requires an image URL and a text prompt. The prompt should describe the desired video motion, which the AI then uses to automate the cinematic transformation.

Does lip-sync video generation work with fal.ai Kling 3.0?

Yes, lip-sync and character scene video generation is handled by fal.ai Kling 3.0. This model supports text-to-video applications and animates characters speaking based on your provided images and prompts.

What inputs are required to create dynamic motion videos from images?

Creating dynamic motion videos requires an image URL and a text prompt describing the desired motion. The AI uses these inputs to automate the cinematic video transformation without manual editing.