ltx2

Generate short AI-driven video clips from text prompts and still images.

1|Updated Apr 11, 2026
One-click install
npx skills add https://github.com/shige1014-dev/backup-OpenMontage --skill ltx2-shige1014-dev
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: ltx2
Source: https://github.com/shige1014-dev/backup-OpenMontage/tree/main/.agents/skills/ltx2
Command: npx skills add https://github.com/shige1014-dev/backup-OpenMontage --skill ltx2-shige1014-dev

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

LTX-2.3 provides a fast path to generate short, high-quality motion clips from text prompts or still images so creators do not need to manually source or animate small footage elements for video projects. It reduces the time and cost of producing b-roll, animated portraits, and motion backgrounds that fit into editing timelines.

Core Features & Use Cases

  • Text-to-video and Image-to-video: Produce ~5 second cinematic clips from descriptive prompts or animate a still image with subtle camera motion.
  • Fine-grained control: Adjustable resolution, frame count, fps, quality mode, and seeds for reproducible outputs; enforces valid frame counts where (n-1) % 8 == 0.
  • Production workflows: Ideal for generating b-roll cutaways, animated slide or portrait backgrounds, branded intro/outro backgrounds, and short motion clips to stitch into Remotion compositions or further post-process with upscaling and audio mixing.
  • Operational notes: Runs on Modal with A100-80GB, requires a Modal LTX2 endpoint URL and a HuggingFace token for weight downloads; expect ~2.5 minutes per default clip and ambient-only generated audio.

Quick Start

Generate a 5-second cinematic b-roll clip of a golden-hour ocean sunset from a text prompt and save it as sunset.mp4.

Frequently Asked Questions about ltx2

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I generate cinematic b-roll from a text prompt for video editing?

You can generate cinematic b-roll by using text-to-video AI to produce ~5 second motion clips from descriptive prompts. This provides fast, high-quality cutaway elements for video production timelines.

Can I animate a still image with subtle camera motion for slide backgrounds?

Yes, you can animate a still image with subtle camera motion using image-to-video generation. This creates short, animated portrait or slide backgrounds ideal for integration into editing pipelines.

Do I need an A100-80GB GPU and Modal endpoint for AI video generation?

Yes, AI video generation requires a Modal LTX2 endpoint URL and an A100-80GB GPU for inference. You also need a HuggingFace token for weight downloads to execute the generation process.

How do I set frame counts for AI video clips to ensure valid output?

To set frame counts for AI video clips, ensure your chosen value follows the rule where (n-1) % 8 == 0. This constraint enforces valid frame counts for correctly generated video outputs.

What are the limitations of generating short AI motion clips for video production?

Limitations of generating short AI motion clips include a ~2.5 minute rendering time per default clip and ambient-only generated audio. Outputs are restricted to ~5 seconds and require significant GPU resources.