ref-to-storyboard

Convert reference images and narrative outlines into structured storyboards and prompts for AI video generation.

Updated May 25, 2026
One-click install
npx skills add https://github.com/fuyucn/skills --skill ref-to-storyboard
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: ref-to-storyboard
Source: https://github.com/fuyucn/skills/tree/main/ref-to-storyboard
Command: npx skills add https://github.com/fuyucn/skills --skill ref-to-storyboard

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) components.

What problem does it solve?

This skill solves the issue of identity drift in AI video generation by bypassing the need for intermediate storyboard images as direct inputs, ensuring consistent facial features from reference photos throughout the entire video production process.

Core Features & Use Cases

  • Identity Preservation: Uses P1/P2 reference photos as direct anchors for Grok video generation to maintain consistent character appearance.
  • Structured Storytelling: Converts story outlines into time-coded, cinematic text scripts optimized for AI video models.
  • Use Case: A creator wants to produce a 10-second romantic scene. They provide two reference photos and a story outline, and the skill generates the precise prompt structure and CLI commands to produce a high-quality, consistent video without the character face-swapping common in standard workflows.

Quick Start

Use the ref-to-storyboard skill to generate a cinematic video script and CLI command for a 10-second scene based on my provided reference photos and story outline.

Frequently Asked Questions about ref-to-storyboard

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I keep character faces consistent in AI video generation from reference photos?

To prevent identity drift in AI video generation, use P1/P2 reference photos as direct anchors for your video model. This bypasses intermediate storyboard images, ensuring facial features remain consistent throughout the entire video production process.

How do I convert a story outline into a cinematic text storyboard?

Converting a narrative outline into a cinematic text storyboard involves generating time-coded scripts optimized for AI video models. This process structures your story into multi-shot sequences with precise prompt mappings for consistent video output.

Do I need the grok-video CLI to generate AI videos from reference images?

Yes, integrating with the grok-video CLI is required to execute the generated prompts and reference image mappings. The skill outputs the structured text scripts and CLI commands needed to produce the final AI video.

What's the best way to create multi-shot AI video sequences without face-swapping?

The best way to avoid face-swapping in multi-shot AI video sequences is bypassing intermediate storyboard images entirely. By using original reference photos as direct inputs for generation, you maintain strict identity preservation across all shots.

Can I use this storyboard prompt engineering approach for short 10-second scenes?

Yes, this approach works for short 10-second scenes by converting your story outline and reference photos into precise prompt structures. It generates the exact CLI commands needed to produce high-quality, consistent cinematic video clips.