talking-head-recut
Package an existing talking-head / interview / podcast video with timed, designed GRAPHIC OVERLAY cards — kinetic titles, lower-thirds, data callouts, quotes, side panels, picture-in-picture — synced to the transcript, on a 16:9 / 9:16 / 4:5 canvas of your choice; the clip plays untouched underneath. Trigger on "graphic overlays", "on-screen graphics", "package / dress up my video". Not plain subtitles (/embedded-captions). Unclear → /hyperframes. Skill: talking-head-recut Owner: heygen-com Summary: Package an existing talking-head / interview / podcast video with timed, designed GRAPHIC OVERLAY cards — kinetic titles, lower-thirds, data callouts, quotes, side panels, picture-in-picture — synced to the transcript, on a 16:9 / 9:16 / 4:5 canvas of your choice; the clip plays untouched underneath. Trigger on "graphic overlays", "on-screen graphics", "package / d
Rank
62
Safety
84
Downloads
1.9k
Updated
Oct 9, 2026
Version
1.0.16
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.9K downloads reported by the source. Last updated 10/9/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 9, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 9, 2026
- Adoption signal
- 1.9K downloadsadoption · observed Oct 9, 2026
- Latest release
- 1.0.16release · observed Oct 4, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s17fpgb0p797dzkbtbrxw5x1hh89qs64:talking-head-recut- Setup complexity is classified as HIGH. You must provision dedicated cloud infrastructure or an isolated VM. Do not run this directly on your local workstation.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-heygen-com-talking-head-recut/snapshot"
Documentation
CLAWHUB
151,177 characters of source documentation, loaded on request.
Extracted files
5 files captured from the source.
SKILL.md
---
name: talking-head-recut
description: Package an existing talking-head / interview / podcast video with timed, designed GRAPHIC OVERLAY cards — kinetic titles, lower-thirds, data callouts, quotes, side panels, picture-in-picture — synced to the transcript, on a 16:9 / 9:16 / 4:5 canvas of your choice; the clip plays untouched underneath. Trigger on "graphic overlays", "on-screen graphics", "package / dress up my video". Not plain subtitles (/embedded-captions). Unclear → /hyperframes.
---
**Plugin installs:** Before setup or freshness commands, follow [plugin execution rules](../hyperframes/references/plugin-installation.md) when this skill is inside a HyperFrames plugin. Standalone installs keep the update instructions below.
> **First, keep this skill fresh — confirm with the user before running:** `npx hyperframes skills update talking-head-recut`. A fast no-op when everything is current; otherwise it refreshes this skill plus the core domain skills it depends on before you rely on them.
# Talking Head Recut
Talking Head Recut takes a local video that **plays in full** and layers a sequence of
timed, designed **graphic cards** onto it — titles, lower-thirds, data callouts,
quotes, side panels, picture-in-picture — synced to what's being said. The agent
designs the cards (timing + content) and **writes each card's HTML directly in the
conversation**, then assembles a single composition HTML and renders it to MP4 via
`hyperframes`. There is no fixed archetype list and no prescribed card structure —
the overlays emerge from what the transcript actually says.
> **The front door is `/hyperframes`.** This skill packages an **existing talking-head clip** with **designed graphic cards** (titles, lower-thirds, data callouts, quotes, side panels, PiP) — not plain captions (the spoken words as text). **The clip plays untouched.** Any other intent — plain subtitles, a standalone graphic, a from-scratch video — or any uncertainty → read `/hyperframes` first: the intent layer owns every route decision.
> **Graphic-packaging sibling of `embedded-captions`.** Captions add the _spoken words_
> as a readable subtitle; this adds _designed graphics_ on top of the playing video.
> Plain subtitles → `embedded-captions`. Build a video from scratch → the creation
> workflows (`product-launch-video` / `faceless-explainer` / …).
Routed through `/hyperframes`, the intent layer confirms only the input (which clip) and **announces** the render-strategy questions as deferred asks — aspect, layout, style group, and card count stay at Step 7, where the probed footage and transcript ground the recommendations; the layer's run-shape questions don't apply. A `BRIEF.md`, when present, carries the confirmed input and any user notes — read it first.
Inspectable intermediate files in the work directory:
- `metadata.json` — duration / width / height / fps
- `audio.mp3` — extracted audio
- `transcript.json` — a flat **word array** `[{ text, start, end }, …]` (Whisper; no_meta.json
{
"ownerId": "kn77d06grj6xqp3dqwkk4bavhn89pegt",
"slug": "talking-head-recut",
"version": "1.0.16",
"publishedAt": 1791142434907
}references/DESIGN_INDEX.md
# V—Take Visual Design Library This directory is a **reference library** for the talking-head-recut skill. Style, layout, and video frame are three **orthogonal** dimensions you can freely mix when designing a takeaway video. ``` Style × Layout × VideoFrame (10) (4) (3) = 120 possible combinations ``` Read a reference file when you decide to use that dimension. Each file is a self-contained HTML fragment that follows the talking-head-recut card-HTML contract (scoped `<style>`, no `<script>`, no external URLs, animations only via `data-anim-*`). ## Layouts — how video and card share the canvas | key | file | what it does | best for | | --------- | -------------------------------------------- | ------------------------------------------------------- | --------------------------------------- | | `split` | [layouts/split.html](layouts/split.html) | 50/50 side-by-side (landscape) or top/bottom (portrait) | speaker + data equal weight | | `stack` | [layouts/stack.html](layouts/stack.html) | video on top (~52%), card below | talking-head with summary card | | `pip` | [layouts/pip.html](layouts/pip.html) | card fills canvas, video rounded PiP in corner | content-heavy moment, speaker secondary | | `overlay` | [layouts/overlay.html](layouts/overlay.html) | video full-bleed, glass card floats on bottom | cinematic / dramatic moments | A layout is a **two-part recipe**: pick a `card.zone` value to put in `storyboard.json` AND author a GSAP tween for `#video-wrap` to its target rect in the composition's `<script>`. Open the layout file's header for the recommended `zone` + the GSAP statement to paste. (Earlier docs referenced a `card.layout` field — that field does NOT exist in the real schema; the strict v3 schema only has `card.zone`.) ## Styles — the card's visual language | key | file | character | accent | suggested font | | ------------ | ------------------------------------------------ | ---------------------------------------------------- | --------- | ------------------- | | `academic` | [styles/academic.html](styles/academic.html) | warm paper · grid · serif · blue highlight | `#2557a7` | serif | | `editorial` | [styles/editorial.html](styles/editorial.html) | cream · coral block · big italic quote | `#ff3a2d` | Playfair-like serif | | `minimal` | [styles/minimal.html](styles/minimal.html) | pure black/white · huge type · generous space | `#000` | Inter | | `spotlight` | [styles/spotlight.html](styles/spotlight.html) | dark purple gradient · glow · dramatic | `#a78bfa` | sans | | `geom`
NOTICE.md
# Attribution The `talking-head-recut` skill (its card-based design system — styles, layouts, frames, fonts, and the GSAP-driven composition workflow) is **adapted from** the open-source **vtake-skills** project (`vtake-cut`): > https://github.com/notedit/vtake-skills Adaptations for this repo: renamed to `talking-head-recut`; transcription repointed to local Whisper via `hyperframes transcribe` (dropping the third-party `@notedit/vtake` CLI and the `vtake.app` proxy); audio/metadata extraction inlined with `ffmpeg`/`ffprobe`; the fixed third-party brand outro removed in favour of an optional, neutral outro; artifacts aligned to the `videos/<project>/` convention. The original is MIT-licensed; its notice is retained below as required. ``` MIT License Copyright (c) 2026 leeoxiang Permission is hereby granted, free of charge, to any person obtaining a copy of this software and associated documentation files (the "Software"), to deal in the Software without restriction, including without limitation the rights to use, copy, modify, merge, publish, distribute, sublicense, and/or sell copies of the Software, and to permit persons to whom the Software is furnished to do so, subject to the following conditions: The above copyright notice and this permission notice shall be included in all copies or substantial portions of the Software. THE SOFTWARE IS PROVIDED "AS IS", WITHOUT WARRANTY OF ANY KIND, EXPRESS OR IMPLIED, INCLUDING BUT NOT LIMITED TO THE WARRANTIES OF MERCHANTABILITY, FITNESS FOR A PARTICULAR PURPOSE AND NONINFRINGEMENT. IN NO EVENT SHALL THE AUTHORS OR COPYRIGHT HOLDERS BE LIABLE FOR ANY CLAIM, DAMAGES OR OTHER LIABILITY, WHETHER IN AN ACTION OF CONTRACT, TORT OR OTHERWISE, ARISING FROM, OUT OF OR IN CONNECTION WITH THE SOFTWARE OR THE USE OR OTHER DEALINGS IN THE SOFTWARE. ```
skill-card.md
## Description: Adds transcript-timed graphic overlays to an existing talking-head video while preserving the underlying clip. This skill is ready for commercial/non-commercial use. ## Publisher: [heygen-com](https://clawhub.ai/user/heygen-com) ### License/Terms of Use: MIT-0 ## Use Case: Video editors and creators use this skill to add designed titles, quotes, data callouts, and other timed graphics to an existing interview, podcast, or talking-head clip. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: Running or updating an unpinned external HyperFrames CLI can change the code the agent executes. Mitigation: Review or pin the HyperFrames version before use, and confirm skill updates before running them. Risk: Extracted audio and transcripts may remain in the project work directory. Mitigation: Avoid sensitive videos unless local storage of their audio and transcripts is acceptable. ## Reference(s): - [Talking Head Recut on ClawHub](https://clawhub.ai/heygen-com/skills/talking-head-recut) - [Visual design reference library](references/DESIGN_INDEX.md) ## Skill Output: **Output Type(s):** [Code, Shell commands, Guidance] **Output Format:** [Markdown guidance with HTML and JSON files, plus a rendered MP4 video] **Output Parameters:** [1D] **Other Properties Related to Output:** [Produces a storyboard, transcript, graphic-card HTML, a composition, and a rendered video.] ## Skill Version(s): 1.0.16 (source: ClawHub release metadata) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/heygen-com/skills/talking-head-recut",
"sourceUrl": "https://clawhub.ai/heygen-com/skills/talking-head-recut",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-09T22:03:15.058Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-heygen-com-talking-head-recut/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-heygen-com-talking-head-recut/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-09T22:03:15.058Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.9K downloads",
"href": "https://clawhub.ai/heygen-com/talking-head-recut",
"sourceUrl": "https://clawhub.ai/heygen-com/talking-head-recut",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-09T22:03:15.058Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.0.16",
"href": "https://clawhub.ai/heygen-com/talking-head-recut",
"sourceUrl": "https://clawhub.ai/heygen-com/talking-head-recut",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-10-04T19:33:54.907Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-heygen-com-talking-head-recut/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-heygen-com-talking-head-recut/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.0.16",
"description": "Synced from 0c3e244 (main)",
"href": "https://clawhub.ai/heygen-com/talking-head-recut",
"sourceUrl": "https://clawhub.ai/heygen-com/talking-head-recut",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-10-04T19:33:54.907Z",
"isPublic": true
}
]
}Record generated Oct 10, 2026.
