抖音爆款爬虫 v2
爬取抖音热榜和搜索建议数据,支持关键词搜索、热榜获取、搜索建议等功能。无需登录即可使用。 Skill: 抖音爆款爬虫 v2 Owner: terrycarter1985 Summary: 爬取抖音热榜和搜索建议数据,支持关键词搜索、热榜获取、搜索建议等功能。无需登录即可使用。 Tags: latest:2.0.1 Version history: v2.0.1 | 2026-06-08T03:49:31.368Z | user v2.0.1: 重写为真实API, 支持热榜/搜索建议/关键词搜索, 无需登录, 支持自然语言入口 v2.0.0 | 2026-06-06T05:04:54.124Z | user 重大更新: 移除prompt注入; 新增自然语言路由; 重写爬虫支持真实抓取+mock降级 v1.3.0 | 2026-06-06T01:49:09.360Z | user 支持自然语言搜索请求;浏览器不可用时自动降级为 mock 数据;修复 Playwright 未安装时的崩溃问题 v1.2.0 | 2026-0
Rank
62
Safety
84
Downloads
1.1k
Updated
Oct 11, 2026
Version
2.0.1
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 11, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 11, 2026
- Adoption signal
- 1.1K downloadsadoption · observed Oct 11, 2026
- Latest release
- 2.0.1release · observed Jun 8, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s17brwfrqyjhbjgadkvar20h8x8492g8:douyin-scraper-v2- Setup complexity is classified as HIGH. You must provision dedicated cloud infrastructure or an isolated VM. Do not run this directly on your local workstation.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-v2/snapshot"
Documentation
CLAWHUB
49,633 characters of source documentation, loaded on request.
Extracted files
5 files captured from the source.
SKILL.md
---
name: douyin-scraper
description: 爬取抖音热榜和搜索建议数据,支持关键词搜索、热榜获取、搜索建议等功能。无需登录即可使用。
version: 2.0.0
---
# 抖音爆款爬虫 Skill
## 功能概述
获取抖音热榜和搜索数据。当前版本使用抖音 Web API,**无需登录**。
## 功能特性
- 🔥 **热榜获取** - 获取当前抖音热搜榜 (无需登录)
- 🔍 **关键词搜索** - 在热榜中匹配关键词 + 获取搜索建议 (无需登录)
- 💡 **搜索建议** - 获取关键词联想 (无需登录)
## ⚠️ 重要说明
抖音搜索 API 需要登录态,当前版本在**无登录**环境下使用以下替代方案:
- 热榜 API:直接获取当前热搜话题
- 搜索建议 API:获取关键词联想
- 搜索时:先在热榜中匹配,再补充搜索建议
如需完整搜索功能,需要提供抖音登录 Cookie。
## 安装
```bash
cd <skill-dir>
npm install
```
Playwright 浏览器(可选,用于完整搜索):
```bash
npx playwright install chromium
```
## 使用方法
### Node.js 版本
```bash
# 搜索关键词
node scripts/douyin_scraper.js search "海鲜" 20
# 获取热榜
node scripts/douyin_scraper.js hot 50
# 获取搜索建议
node scripts/douyin_scraper.js suggest "海鲜售卖"
# 保存到文件
node scripts/douyin_scraper.js search "海鲜" 10 --output result.json
```
### Python 版本
```bash
# 搜索关键词
python scripts/scraper.py search --keyword "海鲜" --limit 20
# 获取热榜
python scripts/scraper.py hot --limit 50
# 获取搜索建议
python scripts/scraper.py suggest --keyword "海鲜售卖"
```
## 自然语言处理指南
当用户用自然语言请求抖音相关数据时,按以下规则解析:
### 搜索意图识别
| 用户说法 | 意图 | 命令 |
|---------|------|------|
| 搜索一下海鲜视频 / 找一些海鲜视频 | 搜索 | `search "海鲜"` |
| 看看抖音热榜 / 抖音最近什么火 | 热榜 | `hot` |
| 海鲜相关的搜索建议 | 建议 | `suggest "海鲜"` |
| 海鲜售卖视频文案 | 搜索 | `search "海鲜售卖"` |
| 分析这个视频链接 xxx | 暂不支持 | 提示用户 |
### 关键词提取
从自然语言中提取关键词:
- "搜索一下**海鲜**视频" → 关键词: `海鲜`
- "找一些**海鲜售卖**相关的视频" → 关键词: `海鲜售卖`
- "**小龙虾**怎么做" → 关键词: `小龙虾`
- "最近**美食**领域什么火" → 关键词: `美食` (搜索热榜匹配)
### 执行流程
1. 解析用户意图 (搜索/热榜/建议)
2. 提取关键词
3. 执行对应命令
4. 格式化展示结果
## 输出数据格式
### 搜索结果
```json
{
"keyword": "海鲜",
"matched_hot": [
{
"rank": 1,
"word": "海鲜话题",
"hot_value": 5000000,
"video_count": 10
}
],
"suggestions": [
{ "word": "海鲜小哥", "group_id": "xxx" }
],
"hot_list": [...],
"note": "在热榜中找到 1 个匹配话题"
}
```
### 热榜数据
```json
[
{
"rank": 1,
"word": "热搜话题",
"hot_value": 12000000,
"video_count": 6,
"group_id": "xxx",
"sentence_id": "xxx"
}
]
```
## 注意事项
1. **请求频率** - 避免频繁调用,建议间隔 >5 秒
2. **数据用途** - 仅供学习和研究
3. **API 限制** - 搜索 API 需登录,热榜和建议 API 无需登录
4. **IP 风控** - 异常请求可能导致 IP 被限
## 故障排除
| 问题 | 解决方案 |
|------|---------|
| API 返回 2483 | 搜索需要登录,使用 `hot` 或 `suggest` 替代 |
| 网络超时 | 检查网络连接,重试 |
| 无匹配结果 | 关键词可能不在热榜,尝试 `suggest` 获取相关词 |README.md
# 抖音爆款爬虫 Skill v2.0 使用抖音 Web API 获取热榜和搜索数据,**无需登录**。 ## 🚀 快速开始 ```bash # 安装 cd douyin-scraper && npm install # 搜索 node scripts/douyin_scraper.js search "海鲜" 20 # 热榜 node scripts/douyin_scraper.js hot 50 # 搜索建议 node scripts/douyin_scraper.js suggest "海鲜售卖" ``` ## 📝 命令 | 命令 | 说明 | 示例 | |------|------|------| | `search <关键词> [数量]` | 搜索 (热榜匹配+建议) | `search "海鲜" 20` | | `hot [数量]` | 获取热榜 | `hot 50` | | `suggest <关键词>` | 搜索建议 | `suggest "海鲜"` | 选项: `--output <文件>` 保存 JSON ## ⚠️ 说明 抖音搜索 API 需登录,本工具在无登录环境下: - 热榜: 直接获取 ✅ - 搜索建议: 直接获取 ✅ - 搜索: 热榜匹配 + 建议补充 ✅ ## 📄 许可 MIT
_meta.json
{
"ownerId": "kn72jmd20ws94jr0p8b3p24wyn82abgj",
"slug": "douyin-scraper-v2",
"version": "2.0.1",
"publishedAt": 1780890571368
}skill-card.md
## Description: 爬取抖音热榜和搜索建议数据,支持关键词搜索、热榜获取、搜索建议等功能。无需登录即可使用。 This skill is for research and development only. ## Publisher: [terrycarter1985](https://clawhub.ai/user/terrycarter1985) ### License/Terms of Use: MIT-0 ## Use Case: Developers and operators use this skill to fetch Douyin hot-search topics, match keywords against the hot list, and retrieve search suggestions for trend monitoring or content research. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: Setup can download mutable Python, Playwright, browser, mirror-hosted, and Docker assets that are broader than the API-only scraper normally needs. Mitigation: Prefer the API-only Node.js or Python scripts when possible, pin and review dependencies before installation, and run browser or Docker setup only in an isolated environment. Risk: Frequent Douyin API requests can fail or trigger IP controls. Mitigation: Throttle usage, follow the artifact guidance to keep requests at least five seconds apart, and stop or retry later if responses indicate rate limiting. Risk: Full Douyin search may require a logged-in browser session or cookie outside the default no-login workflow. Mitigation: Use the hot-list and suggestion modes by default; provide login cookies only when necessary and handle them as sensitive credentials. ## Reference(s): - [ClawHub skill page](https://clawhub.ai/terrycarter1985/skills/douyin-scraper-v2) - [Douyin web service](https://www.douyin.com/) ## Skill Output: **Output Type(s):** [text, markdown, shell commands, configuration, guidance] **Output Format:** [Markdown guidance with shell commands and optional JSON output files] **Output Parameters:** [1D] **Other Properties Related to Output:** [Can save hot-search, search, and suggestion results as JSON when an output file is provided.] ## Skill Version(s): 2.0.1 (source: server release metadata; artifact frontmatter reports 2.0.0) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
examples/search_requests.txt
搜索一下海鲜视频 看看抖音热榜有什么 找一些海鲜售卖相关的视频文案 海鲜相关的搜索建议 最近美食领域什么火 小龙虾怎么做 抖音上什么最火
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/terrycarter1985/skills/douyin-scraper-v2",
"sourceUrl": "https://clawhub.ai/terrycarter1985/skills/douyin-scraper-v2",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T12:50:38.650Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-v2/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-v2/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-11T12:50:38.650Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.1K downloads",
"href": "https://clawhub.ai/terrycarter1985/douyin-scraper-v2",
"sourceUrl": "https://clawhub.ai/terrycarter1985/douyin-scraper-v2",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T12:50:38.650Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "2.0.1",
"href": "https://clawhub.ai/terrycarter1985/douyin-scraper-v2",
"sourceUrl": "https://clawhub.ai/terrycarter1985/douyin-scraper-v2",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-06-08T03:49:31.368Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-v2/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-terrycarter1985-douyin-scraper-v2/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 2.0.1",
"description": "v2.0.1: 重写为真实API, 支持热榜/搜索建议/关键词搜索, 无需登录, 支持自然语言入口",
"href": "https://clawhub.ai/terrycarter1985/douyin-scraper-v2",
"sourceUrl": "https://clawhub.ai/terrycarter1985/douyin-scraper-v2",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-06-08T03:49:31.368Z",
"isPublic": true
}
]
}Record generated Oct 11, 2026.
