smart-short-video

Automate short-video creation by blending AI-generated imagery with original footage.

7|Updated Jan 29, 2026
One-click install
npx skills add https://github.com/temmo1004/smart-short-video --skill smart-short-video
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: smart-short-video
Source: https://github.com/temmo1004/smart-short-video/tree/main
Command: npx skills add https://github.com/temmo1004/smart-short-video --skill smart-short-video

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This Skill automates the end-to-end creation of short-form videos by blending AI-generated imagery with original video segments, reducing manual editing time.

Core Features & Use Cases

  • Automatic video slicing: Splits source video into 3-second segments for modular editing.
  • AI image generation and integration: Generates AI images using multiple services and randomization to accompany scenes.
  • Smart transcription and scripting: Transcribes audio with Whisper and produces concise copy for captions.
  • Scene mixing and pacing: Uses Fisher-Yates shuffle and flexible mix modes to create engaging sequences.
  • End-to-end rendering: Renders a final 9:16 short video via Remotion with optional UI/UX checks.

Quick Start

Run the smart-short-video skill with your source video to generate a 60-second remix using AI images and original clips.

Frequently Asked Questions about smart-short-video

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automate short video creation by blending AI imagery with original footage?

Automating short video creation involves slicing original footage into segments, generating AI images, transcribing audio, and mixing scenes before final rendering. This Skill coordinates these steps end-to-end to produce a 9:16 remix.

Can I use Remotion to render a 9:16 short video with AI-generated images?

Remotion is used to render the final 9:16 short video. The Skill coordinates slicing, transcription, AI-image generation, scene mixing, and final Remotion rendering with guided prompts and safety checks.

How do I transcribe audio and generate captions for a short-form video remix?

Audio is transcribed using Whisper to produce concise copy for captions. This transcription step is integrated into the end-to-end short-video creation workflow.

What is the best way to mix scenes and control pacing for a 60-second video remix?

Scene mixing and pacing utilize a Fisher-Yates shuffle and flexible mix modes to create engaging sequences. You can set target durations of 30, 60, 90, or 120 seconds.

Does this short video generation process support multiple AI image services?

The process supports multiple AI image services and randomization to accompany scenes. It integrates these generated images with original video segments during the scene mixing phase.

What are the limitations of automatically slicing source video for modular editing?

Automatic video slicing splits source video into fixed 3-second segments for modular editing. This rigid segmentation structure may not perfectly align with specific scene transitions or audio cues.