long-video-agent

Automate multi-scene planning, parallel segment generation, and timeline assembly for narrative videos.

12|3|Updated Jun 17, 2026
One-click install
npx skills add https://github.com/phuhao00/bony-agent --skill long-video-agent
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: long-video-agent
Source: https://github.com/phuhao00/bony-agent/tree/main/.agent/skills/long-video-agent
Command: npx skills add https://github.com/phuhao00/bony-agent --skill long-video-agent

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill solves the complexity of producing long-form, multi-scene narrative videos by automating the orchestration of script-to-video workflows, ensuring narrative consistency and rhythmic pacing.

Core Features & Use Cases

  • Multi-Scene Planning: Automatically decomposes scripts into structured shot lists with specific visual descriptions and timing.
  • Parallel Generation: Utilizes advanced models like Tongyi Wan to generate video segments concurrently, significantly reducing production time.
  • Seamless Assembly: Handles the technical stitching of video segments, transitions, and timing to create a cohesive final product.
  • Use Case: A content creator can provide a 2-minute short-story script, and the agent will generate the storyboard, produce individual scenes, and assemble the final video with appropriate pacing.

Quick Start

Use the long-video-agent to generate a 60-second cinematic video based on the provided script about a futuristic city.

Frequently Asked Questions about long-video-agent

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate long-form narrative video production from a script?

To automate long-form narrative video production, this Skill decomposes scripts into structured shot lists, generates segments in parallel, and stitches scenes into a cohesive final video with consistent pacing.

What is multi-scene planning for video generation and how does it work?

Multi-scene planning for video generation is the process of automatically decomposing a script into a structured storyboard with specific visual descriptions and timing, ensuring narrative consistency before parallel segment synthesis begins.

Can I use Tongyi Wan models for parallel video segment generation?

Yes, you can use Tongyi Wan models for parallel video segment generation. The Skill orchestrates concurrent rendering of individual scenes to significantly reduce overall media production time.

How do I seamlessly stitch multiple video segments into a single timeline?

To seamlessly stitch multiple video segments into a single timeline, the Skill handles technical assembly by aligning transitions, timing, and scene pacing to produce a cohesive long-form video.

Does automated scene stitching maintain consistent visual style across shots?

Automated scene stitching maintains consistent visual style across shots by orchestrating multi-scene planning and synthesis parameters, ensuring the final assembled timeline reflects a unified narrative aesthetic.

What are the limitations of AI-driven storyboard generation for complex storytelling?

A limitation of AI-driven storyboard generation for complex storytelling is its reliance on script structure clarity; highly abstract narratives may require manual adjustments to achieve precise rhythmic control and desired visual pacing.