文本识别OCR
兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 Skill: 文本识别OCR Owner: xby-skill Summary: 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 Tags: latest:1.0.0 Version history: v1.0.0 | 2026-07-06T09:23:47.738Z | auto 文本识别OCR v1.0.0 - 新增图像文字识别功能,兼顾速度与精度,支持多种场景(文档、广告牌、截图等)。 - 支持通过图片文件链接或Base64编码输入图像。 - 强制要求配置API密钥,缺失时自动向用户询问。 - 明确工具调用和参数提取流程,提升使用便捷性和安全性。 Archive index: Archive v1.0.0: 8 files, 6843 bytes Files: requirements.txt (78b), scripts/__init__.py (0b), sc
Rank
62
Safety
84
Downloads
1.1k
Updated
Oct 11, 2026
Version
1.0.0
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 11, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 11, 2026
- Adoption signal
- 1.1K downloadsadoption · observed Oct 11, 2026
- Latest release
- 1.0.0release · observed Jul 6, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: medium.
clawhub skill install s173mpn8gcwhe5sh4qmweth8e988w95a:ocr- Python environment detected. Create a strict virtual environment (`python -m venv .venv`) before installing dependencies to prevent system-level package conflicts.
- Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/snapshot"
Documentation
CLAWHUB
4,844 characters of source documentation, loaded on request.
Extracted files
4 files captured from the source.
SKILL.md
--- name: 文本识别OCR description: 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 version: 1.0.0 --- # 文本识别OCR 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 --- ## ⚠️ 强制要求:API 密钥 **此 Skill 必须配置 API 密钥才能使用。** - 首次使用时,如果 `.env` 中没有 `XBY_APIKEY`,**必须使用 AskUserQuestion 工具向用户询问 API 密钥** - 拿到用户提供的密钥后,调用 `scripts.config.set_api_key(api_key)` 保存,然后继续处理 - 获取 API 密钥:https://xiaobenyang.com - **禁止**在缺少 API 密钥时自行搜索或编造数据 --- ## 工作流程(必须遵守) 你(大模型)是路由层,负责理解用户意图、选择工具、提取参数。代码只负责调用API。 ``` 用户输入 → 你选择工具 → 提取该工具需要的参数 → 调用 scripts.tools 中的函数 → 返回结果给用户 ``` ### 步骤 1. **检查 API 密钥**:如果 `scripts.config.settings.api_key` 为空,使用 AskUserQuestion 询问用户,拿到后调用 `scripts.config.set_api_key(key)` 保存 2. **选择工具**:根据用户意图从下方工具列表中选择对应的工具函数 3. **提取参数**:根据选中的工具,提取该工具需要的参数 4. **调用工具**:使用**关键字参数**调用 `scripts.tools` 中的函数,例如 `scripts.tools.search_schools(score='520', province='北京', category='综合')` 5. **返回结果**:将工具返回的 `raw` 数据整理后展示给用户 --- ## 工具选择规则 根据用户意图选择对应的工具函数: | 用户意图 | 工具函数 | |---------|---------| | 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件链接。 | `scripts.tools.ocr` | | 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件的BASE64编码。 | `scripts.tools.ocr_for_data_base64` | **如果参数不完整,使用 AskUserQuestion 向用户询问缺失的参数。** --- ## 工具函数说明 --- ## scripts.tools.ocr 工具描述:兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件链接。 ### 参数定义 |参数名称|参数类型|是否必填|默认值|描述| |------|-------|------|-----|----| |dataUrl|string|true| |图片文件链接地址| --- ## scripts.tools.ocr_for_data_base64 工具描述:兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件的BASE64编码。 ### 参数定义 |参数名称|参数类型|是否必填|默认值|描述| |------|-------|------|-----|----| |dataBase64|string|true| |base64 encoded data of image file| --- --- ## 返回值处理 工具函数返回 `dict` 对象: - `result["raw"]` - API 原始返回数据(JSON),**直接将此数据整理后展示给用户** - `result["success"]` - 是否成功(True/False) - `result["message"]` - 状态消息 --- ## 项目结构 ``` xiaobenyang_gaokao_skill/ ├── scripts/ │ ├── __init__.py │ ├── config.py # 配置管理 + set_api_key() │ ├── call_api.py # API 客户端 + call_api() │ └── tools.py # 工具函数(直接调用) ├── requirements.txt └── SKILL.md ``` --- ## 注意事项 1. **API 密钥是必需的**,无密钥时必须通过 AskUserQuestion 询问用户 2. **禁止**在缺少 API 密钥时自行搜索或编造数据
_meta.json
{
"ownerId": "kn75raw6p9acdmyrzv48w6s3ss88xbca",
"slug": "ocr",
"version": "1.0.0",
"publishedAt": 1783329827738
}skill-card.md
## Description: 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 This skill is ready for commercial/non-commercial use. ## Publisher: [xby-skill](https://clawhub.ai/user/xby-skill) ### License/Terms of Use: MIT-0 ## Use Case: External users and developers can use this skill to send image URLs or base64-encoded images for OCR and receive extracted text for documents, screenshots, signs, and similar image-based text sources. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: Images or base64 image data are sent to XiaoBenYang's remote API for OCR processing. Mitigation: Use the skill only when remote processing is acceptable, and avoid sensitive documents unless the workspace owner has approved that data flow. Risk: The skill asks for an API key and stores it in a local .env file in plaintext. Mitigation: Use a trusted workspace, avoid sharing the project directory or .env file, and rotate the API key if the workspace may have been exposed. ## Reference(s): - [ClawHub skill page](https://clawhub.ai/xby-skill/skills/ocr) - [XiaoBenYang](https://xiaobenyang.com) ## Skill Output: **Output Type(s):** [Text, Markdown, Guidance] **Output Format:** [Markdown or plain text summary of OCR results] **Output Parameters:** [1D] **Other Properties Related to Output:** [The tool response includes success status, raw OCR data, and a status message.] ## Skill Version(s): 1.0.0 (source: frontmatter and release evidence) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
requirements.txt
requests>=2.31.0 pydantic>=2.7.0 pydantic-settings>=2.2.0 python-dotenv>=1.0.1
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/xby-skill/skills/ocr",
"sourceUrl": "https://clawhub.ai/xby-skill/skills/ocr",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T07:01:40.460Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-11T07:01:40.460Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "1.1K downloads",
"href": "https://clawhub.ai/xby-skill/ocr",
"sourceUrl": "https://clawhub.ai/xby-skill/ocr",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-11T07:01:40.460Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.0.0",
"href": "https://clawhub.ai/xby-skill/ocr",
"sourceUrl": "https://clawhub.ai/xby-skill/ocr",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-07-06T09:23:47.738Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.0.0",
"description": "文本识别OCR v1.0.0 - 新增图像文字识别功能,兼顾速度与精度,支持多种场景(文档、广告牌、截图等)。 - 支持通过图片文件链接或Base64编码输入图像。 - 强制要求配置API密钥,缺失时自动向用户询问。 - 明确工具调用和参数提取流程,提升使用便捷性和安全性。",
"href": "https://clawhub.ai/xby-skill/ocr",
"sourceUrl": "https://clawhub.ai/xby-skill/ocr",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-07-06T09:23:47.738Z",
"isPublic": true
}
]
}Record generated Oct 11, 2026.
