talking-head-recut

Add transcript-synced graphic overlay cards to talking-head videos.

255|42|Updated Nov 16, 2023
One-click install
npx skills add https://github.com/chmonitor/chmonitor --skill talking-head-recut
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: talking-head-recut
Source: https://github.com/chmonitor/chmonitor/tree/main/.agents/skills/talking-head-recut
Command: npx skills add https://github.com/chmonitor/chmonitor --skill talking-head-recut

SYSTEM DOCUMENTATION & REQUIREMENTS

💡 This Skill includes references (resource) and assets (resource) components.

What problem does it solve?

This Skill eliminates the tedious, time-consuming work of manually designing, timing, and syncing graphic overlays (titles, lower-thirds, data callouts, quotes, side panels, picture-in-picture) to existing talking-head, interview, or podcast videos, ensuring overlays align perfectly with spoken content without requiring manual video editing or NLE tool use.

Core Features & Use Cases

  • Transcript-synced graphic cards: Automatically generates timed overlay cards synced to local Whisper transcriptions of the video's audio, with no third-party API keys or rate limits required.
  • Customizable visual design: Offers 10 style presets, 4 layout options, and 3 video frame styles to match the video's tone, from academic and editorial to experimental and social media-focused.
  • Use case: Perfect for content creators, podcasters, interview producers, and social media managers who want to add professional, designed on-screen graphics to pre-recorded talking-head footage to increase engagement and clarity.

Quick Start

Use the talking-head-recut skill to add synced graphic overlays to my local talking-head video file at /path/to/my-interview.mp4, using the default auto card count and 16:9 landscape output.

Frequently Asked Questions about talking-head-recut

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I automatically add lower thirds and graphic cards to a talking-head video?

To add lower thirds and graphic cards to a talking-head video, this Skill automatically transcribes your audio with local Whisper and renders timed overlay cards synced directly to the transcript. It outputs a final MP4 without requiring manual video editing.

Does this video packaging tool require third-party API keys for transcript sync?

No, transcript sync for video overlays does not require third-party API keys. The process uses local Whisper transcription via the hyperframes CLI to generate timed graphic cards, avoiding external rate limits and dependencies.

Can I use different visual layouts and style presets for my podcast video overlays?

Yes, you can apply different visual layouts and style presets to your podcast video overlays. The Skill provides 10 style presets, 4 layout options, and 3 video frame styles to match your footage tone, from academic to social media-focused.

What's the best way to sync data callouts and quotes to pre-recorded interview footage?

The best way to sync data callouts and quotes to pre-recorded interview footage is using this Skill's automated transcription matching. It generates timed overlay cards aligned with spoken content, rendering the final graphics directly into the MP4 output.

How do I process a local MP4 file with ffmpeg to add picture-in-picture graphic overlays?

To process a local MP4 file with ffmpeg and add picture-in-picture graphic overlays, the Skill uses ffmpeg and ffprobe for media processing alongside the hyperframes CLI. It automatically handles rendering and timing based on the video's audio transcription.

Will adding transcript-synced graphic cards to my video require re-editing the source footage?

No, adding transcript-synced graphic cards does not require re-editing the source footage. The Skill packages existing talking-head videos by layering designed overlays over the pre-recorded video, leaving the original source file untouched.