抖音爆款爬虫
爬取抖音视频和文案数据。当用户用自然语言请求搜索抖音/抖音视频内容、获取抖音热榜、查找抖音爆款视频时触发。典型触发语包括"搜索一下XX视频"、"帮我找抖音上关于XX的内容"、"抖音上有什么XX"、"抖音热榜"、"抖音热门"等。支持中文关键词搜索和热榜获取。 Skill: 抖音爆款爬虫 Owner: terrycarter1985 Summary: 爬取抖音视频和文案数据。当用户用自然语言请求搜索抖音/抖音视频内容、获取抖音热榜、查找抖音爆款视频时触发。典型触发语包括"搜索一下XX视频"、"帮我找抖音上关于XX的内容"、"抖音上有什么XX"、"抖音热榜"、"抖音热门"等。支持中文关键词搜索和热榜获取。 Tags: latest:1.2.0 Version history: v1.2.0 | 2026-05-31T04:47:28.252Z | user 支持自然语言触发搜索,浏览器不可用时自动降级为模拟数据 v1.1.0 | 2026-05-20T00:03:50.737Z | user 支持真实搜索数据(移动端UA),自然语言关键词解析,venv依赖安装 Archive index: Archive v1.2.0: 10 files, 11822 bytes Files: inst
Rank
62
Safety
84
Downloads
1.1k
Updated
Oct 11, 2026
Version
1.2.0
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 11, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 11, 2026
- Adoption signal
- 1.1K downloadsadoption · observed Oct 11, 2026
- Latest release
- 1.2.0release · observed May 31, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: medium.
clawhub skill install s17brwfrqyjhbjgadkvar20h8x8492g8:douyin-scraper-terrycarter- Python environment detected. Create a strict virtual environment (`python -m venv .venv`) before installing dependencies to prevent system-level package conflicts.
- Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-terrycarter/snapshot"
Documentation
CLAWHUB
15,642 characters of source documentation, loaded on request.
Extracted files
5 files captured from the source.
SKILL.md
--- name: douyin-scraper description: 爬取抖音视频和文案数据。当用户用自然语言请求搜索抖音/抖音视频内容、获取抖音热榜、查找抖音爆款视频时触发。典型触发语包括"搜索一下XX视频"、"帮我找抖音上关于XX的内容"、"抖音上有什么XX"、"抖音热榜"、"抖音热门"等。支持中文关键词搜索和热榜获取。 --- # 抖音爆款爬虫 ## 用法 当用户用自然语言请求搜索抖音内容时,提取关键词后调用脚本: ### 搜索视频 ```bash python3 scripts/scraper.py search --keyword "<关键词>" --limit <数量> ``` ### 获取热榜 ```bash python3 scripts/scraper.py search --keyword "<分类>" --limit <数量> # 或 python3 scripts/scraper.py hot --category "<分类>" --limit <数量> ``` ### 保存结果 加 `--output <文件名>` 和 `--format json|csv` 保存到文件。 ## 自然语言 → 命令映射 | 用户说 | 命令 | |--------|------| | 搜索一下海鲜视频 | `python3 scripts/scraper.py search --keyword "海鲜" --limit 10` | | 帮我找抖音上关于小龙虾的内容 | `python3 scripts/scraper.py search --keyword "小龙虾" --limit 10` | | 抖音热榜 | `python3 scripts/scraper.py hot --limit 20` | | 美食热门视频 | `python3 scripts/scraper.py hot --category "美食" --limit 20` | ## 输出格式 每条结果包含:title, description, author, play_count, like_count, comment_count, share_count, url, tags, publish_time。 ## 注意事项 - 浏览器不可用时自动降级为模拟数据(会打印提示) - 遵守平台规则,避免频繁请求 - 仅供学习研究使用
README.md
# 抖音爆款爬虫 Skill
使用 Playwright 自动化浏览器操作,爬取抖音爆款视频和文案数据。
## 📦 安装
### 方式一:一键安装(推荐)
```bash
# 进入 skill 目录
cd /root/.openclaw/workspace/skills/douyin-scraper
# 运行安装脚本
./install.sh
```
### 方式二:手动安装 - Python 版本
```bash
# 进入 skill 目录
cd /root/.openclaw/workspace/skills/douyin-scraper
# 创建虚拟环境
python3 -m venv venv
# 激活虚拟环境
source venv/bin/activate
# 安装依赖
pip install playwright
# 安装浏览器
playwright install chromium
```
### 方式三:Node.js 版本
```bash
# 进入 skill 目录
cd /root/.openclaw/workspace/skills/douyin-scraper
# 安装依赖
npm install
# 安装浏览器
npx playwright install chromium
```
## 🚀 快速开始
### 方式一:使用启动脚本(推荐)
```bash
# 搜索关键词
./run.sh search --keyword "海鲜" --limit 10
# 获取热榜
./run.sh hot --limit 20
# 搜索并保存结果
./run.sh search --keyword "海鲜售卖" --limit 20 --output seafood_videos.json
```
### 方式二:手动激活虚拟环境
```bash
# 激活虚拟环境
source venv/bin/activate
# 搜索关键词
python scripts/scraper.py search --keyword "海鲜" --limit 10
# 获取热榜
python scripts/scraper.py hot --limit 20
# 搜索并保存结果
python scripts/scraper.py search --keyword "海鲜售卖" --limit 20 --output seafood_videos.json
```
### Node.js 版本
```bash
# 搜索关键词
node scripts/douyin_scraper.js search "海鲜" 10
# 获取热榜
node scripts/douyin_scraper.js hot 20
# 搜索并保存结果
node scripts/douyin_scraper.js search "海鲜售卖" 20 seafood_sales.json json
```
## 📝 使用示例
### 示例 1:搜索海鲜售卖视频
```bash
# Python 版本
python scripts/scraper.py search --keyword "海鲜售卖" --limit 15 --output seafood_sales.json
# Node.js 版本
node scripts/douyin_scraper.js search "海鲜售卖" 15 seafood_sales.json json
```
### 示例 2:获取美食热榜
```bash
# Python 版本
python scripts/scraper.py hot --category "美食" --limit 20 --output food_hot.json
# Node.js 版本
node scripts/douyin_scraper.js hot "美食" 20 food_hot.json json
```
### 示例 3:导出 CSV 格式
```bash
# Python 版本
python scripts/scraper.py search --keyword "小龙虾" --limit 10 --format csv --output crayfish.csv
# Node.js 版本
node scripts/douyin_scraper.js search "小龙虾" 10 crayfish.csv csv
```
## 📊 输出数据格式
### JSON 格式
```json
[
{
"title": "视频标题",
"description": "视频描述",
"author": "作者昵称",
"play_count": 1000000,
"like_count": 50000,
"comment_count": 2000,
"share_count": 1000,
"url": "https://www.douyin.com/video/xxx",
"tags": ["标签1", "标签2"],
"publish_time": "2026-03-21"
}
]
```
### CSV 格式
```csv
title,author,play_count,like_count,comment_count,url,tags
视频标题,作者昵称,1000000,50000,2000,https://...,标签1|标签2
```
## ⚙️ 配置选项
### Python 版本
```bash
# 无头模式(不显示浏览器)
--headless
# 请求间隔(秒)
--delay 3.0
# 输出格式
--format json|csv
```
### Node.js 版本
在代码中修改配置:
```javascript
const scraper = new DouyinScraper({
headless: true, // 无头模式
delay: 2000 // 请求间隔(毫秒)
});
```
## ⚠️ 注意事项
1. **遵守抖音平台规则** - 合理使用,避免频繁请求
2. **请求间隔** - 建议在请求之间添加适当延时(默认 2 秒)
3. **数据用途** - 仅供学习和研究使用
4. **账号安全** - 不要登录账号,避免风控
5. **IP 限制** - 注意 IP 被封禁的风险
## 🔧 故障排除
### 问题:浏览器启动失败
**解决方案:**
```bash
# Python 版本
playwright install chromium
# Node.js 版本
npx playwright install chromium
```
### 问题:页面加载超时
**解决方案:**
- 增加超时时间
- 检查网络连接
-_meta.json
{
"ownerId": "kn72jmd20ws94jr0p8b3p24wyn82abgj",
"slug": "douyin-scraper-terrycarter",
"version": "1.2.0",
"publishedAt": 1780202848252
}skill-card.md
## Description: 为代理提供抖音关键词搜索和热榜查询命令,可输出视频标题、作者、互动计数、链接和标签字段;当前实现可能返回模拟示例数据。 This skill is for research and development only. ## Publisher: [terrycarter1985](https://clawhub.ai/user/terrycarter1985) ### License/Terms of Use: MIT-0 ## Use Case: Developers and external users can ask an agent in Chinese to run Douyin keyword search or hot-list collection workflows and save video-style result records for analysis. Treat outputs as example data unless the publisher verifies real extraction. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: The skill advertises Douyin scraping while current search and hot-list results may be synthetic example data. Mitigation: Treat results as examples unless the publisher verifies extraction, and disclose uncertainty when using the data. Risk: The skill may contact Douyin, install browser automation components, and create local output files. Mitigation: Run it only in an isolated environment, review install scripts before execution, avoid logging into accounts, and comply with platform rules. Risk: Browser setup includes risky automation choices that should be contained before installation. Mitigation: Prefer a sandboxed container or other restricted runtime and review browser launch flags before use. ## Reference(s): - [ClawHub skill page](https://clawhub.ai/terrycarter1985/skills/douyin-scraper-terrycarter) - [Playwright documentation](https://playwright.dev/) ## Skill Output: **Output Type(s):** [text, shell commands, JSON, CSV files, guidance] **Output Format:** [Console text and optional JSON or CSV output files] **Output Parameters:** [1D] **Other Properties Related to Output:** [Records include title, description, author, play_count, like_count, comment_count, share_count, url, tags, and publish_time; browser fallback can produce synthetic examples.] ## Skill Version(s): 1.2.0 (source: server release metadata) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
requirements.txt
playwright>=1.40.0
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/terrycarter1985/skills/douyin-scraper-terrycarter",
"sourceUrl": "https://clawhub.ai/terrycarter1985/skills/douyin-scraper-terrycarter",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T08:28:15.866Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-terrycarter/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-terrycarter/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-11T08:28:15.866Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.1K downloads",
"href": "https://clawhub.ai/terrycarter1985/douyin-scraper-terrycarter",
"sourceUrl": "https://clawhub.ai/terrycarter1985/douyin-scraper-terrycarter",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T08:28:15.866Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.2.0",
"href": "https://clawhub.ai/terrycarter1985/douyin-scraper-terrycarter",
"sourceUrl": "https://clawhub.ai/terrycarter1985/douyin-scraper-terrycarter",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-05-31T04:47:28.252Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-terrycarter/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-terrycarter/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.2.0",
"description": "支持自然语言触发搜索,浏览器不可用时自动降级为模拟数据",
"href": "https://clawhub.ai/terrycarter1985/douyin-scraper-terrycarter",
"sourceUrl": "https://clawhub.ai/terrycarter1985/douyin-scraper-terrycarter",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-05-31T04:47:28.252Z",
"isPublic": true
}
]
}Record generated Oct 11, 2026.
