talking-head-recut

Package talking-head videos with timed graphic overlays using Whisper transcription and hyperframes rendering.

Updated Mar 3, 2026
One-click install
npx skills add https://github.com/X-RANKFLOW-MEDIA-GROUP/masseurmatch --skill talking-head-recut-x-rankflow-media-group
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: talking-head-recut
Source: https://github.com/X-RANKFLOW-MEDIA-GROUP/masseurmatch/tree/main/.agents/skills/talking-head-recut
Command: npx skills add https://github.com/X-RANKFLOW-MEDIA-GROUP/masseurmatch --skill talking-head-recut-x-rankflow-media-group

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill requires ffmpeg, ffprobe, hyperframes, and includes scripts (resource) and references (resource) and assets (resource) components.

What problem does it solve?

This Skill allows you to design and package existing talking-head videos with timed graphic overlays, enhancing the video content with additional information, graphics, and branding.

Core Features & Use Cases

  • Design Graphic Overlays: Create and apply timed, designed graphic overlays to videos like titles, lower-thirds, data callouts, quotes, side panels, and picture-in-picture.
  • Transcription and Syncing: Extract audio from videos and transcribe it for accurate overlay timing.
  • Customization: Customize the visual style, layout, and card count based on the video content and user preferences.
  • Use Case: Use this Skill to package a customer testimonial video with branded graphics and relevant data points overlaid in real-time, creating a more engaging viewing experience.

Quick Start

Use the talking-head-recut skill to package the video 'customer-testimonial.mp4' with a custom overlay design, choosing the '4:5' aspect ratio, 'stack' layout, 'editorial' style, and 'Auto' card count.

Frequently Asked Questions about talking-head-recut

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I add timed graphic overlays to a talking-head video?

To add timed graphic overlays to a talking-head video, this Skill transcribes the audio with Whisper and uses hyperframes to render titles, lower-thirds, and data callouts in sync with the spoken content. It packages the video with customized branding and layout.

What is the best way to package video content with transcription-synced graphics?

Packaging video content with transcription-synced graphics is best handled by extracting audio for transcription, then using the timestamps to trigger visual styles like editorial layouts or stacked side panels. This approach automatically aligns data callouts with spoken words.

Do I need ffmpeg and hyperframes to design video overlays with this tool?

Yes, you need ffmpeg, ffprobe, and the hyperframes CLI installed to use this tool for designing video overlays. These dependencies handle audio extraction, media processing, and the final rendering of the timed graphic overlays onto your video.

Can I customize the aspect ratio and layout style for video packaging?

Yes, you can customize the aspect ratio to formats like 4:5 and select layout styles such as stack or editorial for your video packaging. The Skill adjusts the visual style, layout, and card count based on your specific content preferences.

What types of graphic overlays can I apply to testimonial videos?

You can apply various graphic overlays to testimonial videos, including titles, lower-thirds, data callouts, quotes, side panels, and picture-in-picture elements. These designed overlays enhance the viewing experience by adding relevant data points in real-time.

Why use Whisper transcription for video editing and overlay timing?

Using Whisper transcription for video editing ensures accurate overlay timing by aligning graphic appearances with the exact spoken words. This automated syncing process prevents manual timeline adjustments and keeps visual elements contextually relevant to the audio.