视频对口型 Video Retalk
Tongyi VideoRetalk lip sync / lip-sync (mouth sync, dubbing) video model — takes a talking-person video plus a voice audio track and regenerates the video so the speaker's mouth/lips match the new audio. Use this for lip syncing a person video to new speech. Optionally provide a reference face image to pick the target person when the video contains multiple faces. 通义声动人像 VideoRetalk 口型同步(对口型、lip sync / lip-sync、配音对嘴)视频模型,输入一段人物讲话视频与一段人声音频,生成讲话口型与音频匹配的新视频;适用于让人物视频的口型对上新的语音。当视频中存在多张人脸时,可额外提供人脸参考图来指定要替换口型的目标人物。
Rank
62
Safety
84
Downloads
3.0k
Updated
Oct 9, 2026
Version
1.3.26
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 3K downloads reported by the source. Last updated 10/9/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 9, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 9, 2026
- Adoption signal
- 3K downloadsadoption · observed Oct 9, 2026
- Latest release
- 1.3.26release · observed Oct 8, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s170j1ymymrxasgd00dsk7tckx84cf45:dlazy-videoretalk- Install using `clawhub skill install s170j1ymymrxasgd00dsk7tckx84cf45:dlazy-videoretalk` in an isolated environment before connecting it to live workloads.
- No published capability contract is available yet, so validate auth and request/response behavior manually.
- Review the upstream CLAWHUB listing at https://clawhub.ai/dlazyai/dlazy-videoretalk before using production credentials.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-dlazyai-dlazy-videoretalk/snapshot"
Documentation
CLAWHUB
146,359 characters of source documentation, loaded on request.
Extracted files
4 files captured from the source.
SKILL.md
---
name: dlazy-videoretalk
version: 1.3.5
description: "Tongyi VideoRetalk lip sync / lip-sync (mouth sync, dubbing) video model — takes a talking-person video plus a voice audio track and regenerates the video so the speaker's mouth/lips match the new audio. Use this for lip syncing a person video to new speech. Optionally provide a reference face image to pick the target person when the video contains multiple faces. 通义声动人像 VideoRetalk 口型同步(对口型、lip sync / lip-sync、配音对嘴)视频模型,输入一段人物讲话视频与一段人声音频,生成讲话口型与音频匹配的新视频;适用于让人物视频的口型对上新的语音。当视频中存在多张人脸时,可额外提供人脸参考图来指定要替换口型的目标人物。"
metadata: {"clawdbot":{"emoji":"🤖","requires":{"bins":["npm","npx"]},"install":"npm install -g @dlazy/[email protected]","installAlternative":"npx @dlazy/[email protected]","homepage":"https://github.com/dlazy-ai/cli","source":"https://github.com/dlazy-ai/cli","author":"dlazyai","license":"see-repo","npm":"https://www.npmjs.com/package/@dlazy/cli","configLocation":"~/.dlazy/config.json","apiEndpoints":["api.dlazy.com","files.dlazy.com"]},"openclaw":{"systemPrompt":"When invoking this skill, use dlazy videoretalk -h for help."}}
---
# 视频对口型 Video Retalk
[English](./SKILL.md) · [中文](./SKILL-cn.md)
Tongyi VideoRetalk lip sync / lip-sync (mouth sync, dubbing) video model — takes a talking-person video plus a voice audio track and regenerates the video so the speaker's mouth/lips match the new audio. Use this for lip syncing a person video to new speech. Optionally provide a reference face image to pick the target person when the video contains multiple faces.
## Trigger Keywords
- videoretalk
## Authentication
All requests require a dLazy API key. The recommended way to authenticate is:
```bash
dlazy login
```
This runs a device-code flow (also works in remote shells) and **automatically saves your API key** to the local CLI config — no manual copy/paste required.
### Alternative: Set the Key Manually
If you already have an API key, you can save it directly:
```bash
dlazy auth set YOUR_API_KEY
```
The CLI saves the key in your user config directory (`~/.dlazy/config.json` on macOS/Linux, `%USERPROFILE%\.dlazy\config.json` on Windows), with file permissions restricted to your OS user account. You can also supply the key per-invocation via the `DLAZY_API_KEY` environment variable.
### Getting Your API Key Manually
1. Sign in or create an account at [dlazy.com](https://dlazy.com)
2. Go to [dlazy.com/dashboard/organization/api-key](https://dlazy.com/dashboard/organization/api-key)
3. Copy the key shown in the API Key section
Each key is scoped to your dLazy organization and can be **rotated or revoked at any time** from the same dashboard.
## About & Provenance
- **CLI source code**: [github.com/dlazy-ai/cli](https://github.com/dlazy-ai/cli)
- **Maintainer**: dlazyai
- **npm package**: `@dlazy/cli` (pinned to `1.2.3` in this skill's install spec)
- **Homepage**: [dlazy.com](https://dlazy.com)
You can install on demand without persisting_meta.json
{
"ownerId": "kn7c5wgeajfcfvdfb5ceemvdb984cjpd",
"slug": "dlazy-videoretalk",
"version": "1.3.26",
"publishedAt": 1791423342707
}skill-card.md
## Description: Helps agents create a lip-synced talking-person video from a source video and voice audio, with an optional reference face for multi-person footage. This skill is ready for commercial/non-commercial use. ## Publisher: [dlazyai](https://clawhub.ai/user/dlazyai) ### License/Terms of Use: MIT-0 ## Use Case: Creators and developers use this skill to align a person's mouth movements in a video with new speech audio, optionally selecting a face in multi-person footage. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: Video, audio, and optional face images may be uploaded to dLazy's hosted service. Mitigation: Share only media you are authorized to process and review dLazy's service terms before submitting sensitive content. Risk: Logging in can save a dLazy API key locally, and installing the third-party CLI runs external software. Mitigation: Protect the saved key and review the CLI before use; prefer the pinned on-demand npx command to avoid a global installation. Risk: The skill's example command and output sample do not match its documented video and audio inputs. Mitigation: Check the installed CLI's videoretalk help and validate its actual result before relying on the examples. ## Reference(s): - [ClawHub skill listing](https://clawhub.ai/dlazyai/skills/dlazy-videoretalk) - [dLazy CLI homepage](https://github.com/dlazy-ai/cli) - [dLazy CLI npm package](https://www.npmjs.com/package/@dlazy/cli) ## Skill Output: **Output Type(s):** [Guidance, Shell commands, JSON response] **Output Format:** [Text guidance and CLI JSON containing a generated-video URL; optional downloaded video file] **Output Parameters:** [1D] **Other Properties Related to Output:** [Asynchronous requests may return a task ID instead of a completed result.] ## Skill Version(s): 1.3.26 (source: ClawHub release evidence; artifact frontmatter says 1.3.5) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
SKILL-cn.md
---
name: dlazy-videoretalk
version: 1.3.5
description: "通义声动人像 VideoRetalk 口型同步(对口型、lip sync / lip-sync、配音对嘴)视频模型,输入一段人物讲话视频与一段人声音频,生成讲话口型与音频匹配的新视频;适用于让人物视频的口型对上新的语音。当视频中存在多张人脸时,可额外提供人脸参考图来指定要替换口型的目标人物。"
metadata: {"clawdbot":{"emoji":"🤖","requires":{"bins":["npm","npx"]},"install":"npm install -g @dlazy/[email protected]","installAlternative":"npx @dlazy/[email protected]","homepage":"https://github.com/dlazy-ai/cli","source":"https://github.com/dlazy-ai/cli","author":"dlazyai","license":"see-repo","npm":"https://www.npmjs.com/package/@dlazy/cli","configLocation":"~/.dlazy/config.json","apiEndpoints":["api.dlazy.com","files.dlazy.com"]},"openclaw":{"systemPrompt":"当调用此技能时,可以使用 dlazy videoretalk -h 查看帮助信息。"}}
---
# 视频对口型 Video Retalk
[English](./SKILL.md) · [中文](./SKILL-cn.md)
通义声动人像 VideoRetalk 口型同步(对口型、lip sync / lip-sync、配音对嘴)视频模型,输入一段人物讲话视频与一段人声音频,生成讲话口型与音频匹配的新视频;适用于让人物视频的口型对上新的语音。当视频中存在多张人脸时,可额外提供人脸参考图来指定要替换口型的目标人物。
## 触发关键词
- videoretalk
## 身份验证 (Authentication)
所有请求都需要 dLazy API key。**推荐使用** `dlazy login` 完成登录:
```bash
dlazy login
```
该命令使用设备码流程(远程终端也可用),登录成功后 **自动把 API key 写入本地 CLI 配置**,无需手动复制粘贴。
### 备选:手动设置 API Key
如果你已有 API key,也可以直接保存:
```bash
dlazy auth set YOUR_API_KEY
```
CLI 会把 key 保存在你的用户配置目录(macOS/Linux 上为 `~/.dlazy/config.json`,Windows 上为 `%USERPROFILE%\.dlazy\config.json`),文件权限仅限当前操作系统用户访问。你也可以用 `DLAZY_API_KEY` 环境变量按次传入。
### 手动获取 API Key
1. 登录或在 [dlazy.com](https://dlazy.com) 创建账号
2. 访问 [dlazy.com/dashboard/organization/api-key](https://dlazy.com/dashboard/organization/api-key)
3. 复制 API Key 区域显示的密钥
每个 key 都属于你自己的 dLazy 组织,可在同一控制面板**随时轮换或吊销**。
## 关于与来源 (Provenance)
- **CLI 源代码**: [github.com/dlazy-ai/cli](https://github.com/dlazy-ai/cli)
- **维护者**: dlazyai
- **npm 包名**: `@dlazy/cli`(本技能 install 字段固定到 `1.2.3` 版本)
- **官网**: [dlazy.com](https://dlazy.com)
如果你不希望在系统上长期保留一个全局 CLI,可以按需运行:
```bash
npx @dlazy/[email protected] <command>
```
如选择全局安装,技能的 `metadata.clawdbot.install` 字段已固定到 `npm install -g @dlazy/[email protected]`。安装前建议先到 GitHub 仓库审阅源码。
## 工作原理
此技能是 dLazy 托管 API 的轻量封装。调用时:
- 你提供的提示词与参数会发送到 dLazy API(`api.dlazy.com`)进行推理。
- 传入图像 / 视频 / 音频字段的本地文件路径会被 CLI 上传到 dLazy 媒体存储(`files.dlazy.com`),以便模型读取 —— 与任何云端生成 API 的流程一致。
- API 返回的生成结果 URL 由 `files.dlazy.com` 托管。
这是标准的 SaaS 调用模式;技能本身不会越权访问网络或文件系统,所有动作都由 dLazy CLI 完成。完整服务条款请参见 [dlazy.com](https://dlazy.com)。
## 使用方法
**CRITICAL INSTRUCTION FOR AGENT**:
执行 `dlazy videoretalk` 命令获取结果。
```bash
dlazy videoretalk -h
Options:
--video_url [video_url] 视频链接 [video: url or local path]
--audio_url [audio_url] 音频链接 [audio: url or local path]
--ref_image_url [ref_image_url] 人脸参考图 [image: url or local path]
--video_extension [video_extension] 按音频时长扩展视频 [default: false] (choices: "true", "false")
--query_face_threshold [query_face_threshold]人脸匹配置信度 [default: 170] [only when ref_image_url non-empty]
--dry-run Print payload + cost estimate withoAionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/dlazyai/skills/dlazy-videoretalk",
"sourceUrl": "https://clawhub.ai/dlazyai/skills/dlazy-videoretalk",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-09T10:19:11.228Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-dlazyai-dlazy-videoretalk/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-dlazyai-dlazy-videoretalk/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-09T10:19:11.228Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "3K downloads",
"href": "https://clawhub.ai/dlazyai/dlazy-videoretalk",
"sourceUrl": "https://clawhub.ai/dlazyai/dlazy-videoretalk",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-09T10:19:11.228Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.3.26",
"href": "https://clawhub.ai/dlazyai/dlazy-videoretalk",
"sourceUrl": "https://clawhub.ai/dlazyai/dlazy-videoretalk",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-10-08T01:35:42.707Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-dlazyai-dlazy-videoretalk/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-dlazyai-dlazy-videoretalk/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.3.26",
"description": "例行版本更新 2026-10-08",
"href": "https://clawhub.ai/dlazyai/dlazy-videoretalk",
"sourceUrl": "https://clawhub.ai/dlazyai/dlazy-videoretalk",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-10-08T01:35:42.707Z",
"isPublic": true
}
]
}Record generated Oct 9, 2026.
