comfyui-video-pipeline

Route prompts and assets to suitable engines for ComfyUI video generation workflows.

86|24|Updated Feb 6, 2026
One-click install
npx skills add https://github.com/MCKRUZ/ComfyUI-Expert --skill comfyui-video-pipeline-mckruz
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: comfyui-video-pipeline
Source: https://github.com/MCKRUZ/ComfyUI-Expert/tree/main/skills/comfyui-video-pipeline
Command: npx skills add https://github.com/MCKRUZ/ComfyUI-Expert --skill comfyui-video-pipeline-mckruz

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Orchestrates end-to-end ComfyUI video generation workflows by routing prompts and assets to the most suitable engine.

Core Features & Use Cases

  • Automated engine selection based on quality, duration, and VRAM constraints.
  • Supports Wan 2.2 MoE, FramePack, and AnimateDiff pipelines for image-to-video, text-to-video, and motion-controlled animation.
  • Talking head workflows with end-to-end lip-sync, enhancement, and post-processing options.

Quick Start

Provide your image or text prompts and let the pipeline automatically select the best engine for your delivery requirements.

Frequently Asked Questions about comfyui-video-pipeline

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I orchestrate ComfyUI video generation pipelines across different engines?

To orchestrate ComfyUI video generation, this pipeline routes your text prompts and image assets to the most suitable engine, automating workflow assembly for Wan 2.2 MoE, FramePack, and AnimateDiff.

How does ComfyUI handle automated engine selection for AI video generation?

Automated engine selection for AI video generation works by evaluating your quality, duration, and VRAM constraints, then routing the inputs to the appropriate ComfyUI pipeline like Wan 2.2 MoE or FramePack.

Can I create talking head videos with lip-sync in ComfyUI?

Yes, you can create talking head videos in ComfyUI using dedicated workflows that provide end-to-end lip-sync, enhancement, and post-processing options for your input assets.

What is the best way to convert image-to-video or text-to-video in ComfyUI?

The best way to convert image-to-video or text-to-video is using an orchestrated pipeline that automatically applies frame and quality settings, post-processing, and output assembly across supported engines.

Do I need high VRAM to run Wan 2.2 MoE and FramePack pipelines?

VRAM requirements vary by engine, but the pipeline includes resource awareness to help select an engine like FramePack or Wan 2.2 MoE that best fits your available hardware constraints.

Why use an orchestrated pipeline instead of manually building ComfyUI video workflows?

Using an orchestrated pipeline prevents manual routing errors by automatically handling engine selection, frame settings, and post-processing, satisfying end-to-end video generation requirements efficiently.