agentCLAWHUBUnverified

文本识别OCR

兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 Skill: 文本识别OCR Owner: xby-skill Summary: 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 Tags: latest:1.0.0 Version history: v1.0.0 | 2026-07-06T09:23:47.738Z | auto 文本识别OCR v1.0.0 - 新增图像文字识别功能,兼顾速度与精度,支持多种场景(文档、广告牌、截图等)。 - 支持通过图片文件链接或Base64编码输入图像。 - 强制要求配置API密钥,缺失时自动向用户询问。 - 明确工具调用和参数提取流程,提升使用便捷性和安全性。 Archive index: Archive v1.0.0: 8 files, 6843 bytes Files: requirements.txt (78b), scripts/__init__.py (0b), sc

OpenClaw

Rank

62

Safety

84

Downloads

1.1k

Updated

Oct 11, 2026

Version

1.0.0

Source

CLAWHUB

About

What it does, and when to use it.

Capability contract not published. No trust telemetry is available yet. 1.1K downloads reported by the source. Last updated 10/11/2026.

Avoid when

  • Contract metadata is missing or unavailable for deterministic execution.

Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing

Public facts

Every fact links back to the source it came from.

Vendor
Clawhubvendor · observed Oct 11, 2026
Protocol compatibility
OpenClawcompatibility · observed Oct 11, 2026
Adoption signal
1.1K downloadsadoption · observed Oct 11, 2026
Latest release
1.0.0release · observed Jul 6, 2026
Handshake status
UNKNOWNsecurity

Install and run

Setup complexity: medium.

clawhub skill install s173mpn8gcwhe5sh4qmweth8e988w95a:ocr
  1. Python environment detected. Create a strict virtual environment (`python -m venv .venv`) before installing dependencies to prevent system-level package conflicts.
  2. Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
  3. Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.

Contract: missing

curl -s "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/snapshot"

Documentation

CLAWHUB

4,844 characters of source documentation, loaded on request.

Extracted files

4 files captured from the source.

SKILL.md

---
name: 文本识别OCR
description: 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。
version: 1.0.0
---

# 文本识别OCR

兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。

---

## ⚠️ 强制要求:API 密钥

**此 Skill 必须配置 API 密钥才能使用。**

- 首次使用时,如果 `.env` 中没有 `XBY_APIKEY`,**必须使用 AskUserQuestion 工具向用户询问 API 密钥**
- 拿到用户提供的密钥后,调用 `scripts.config.set_api_key(api_key)` 保存,然后继续处理
- 获取 API 密钥:https://xiaobenyang.com
- **禁止**在缺少 API 密钥时自行搜索或编造数据

---

## 工作流程(必须遵守)

你(大模型)是路由层,负责理解用户意图、选择工具、提取参数。代码只负责调用API。

```
用户输入 → 你选择工具 → 提取该工具需要的参数 → 调用 scripts.tools 中的函数 → 返回结果给用户
```

### 步骤

1. **检查 API 密钥**:如果 `scripts.config.settings.api_key` 为空,使用 AskUserQuestion 询问用户,拿到后调用 `scripts.config.set_api_key(key)` 保存
2. **选择工具**:根据用户意图从下方工具列表中选择对应的工具函数
3. **提取参数**:根据选中的工具,提取该工具需要的参数
4. **调用工具**:使用**关键字参数**调用 `scripts.tools` 中的函数,例如 `scripts.tools.search_schools(score='520', province='北京', category='综合')`
5. **返回结果**:将工具返回的 `raw` 数据整理后展示给用户

---
## 工具选择规则

根据用户意图选择对应的工具函数:

| 用户意图 | 工具函数 | 
|---------|---------|
| 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件链接。 | `scripts.tools.ocr` |
| 兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件的BASE64编码。 | `scripts.tools.ocr_for_data_base64` |

**如果参数不完整,使用 AskUserQuestion 向用户询问缺失的参数。**

---

## 工具函数说明

---

## scripts.tools.ocr
工具描述:兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件链接。
### 参数定义
|参数名称|参数类型|是否必填|默认值|描述|
|------|-------|------|-----|----|
|dataUrl|string|true| |图片文件链接地址|

---

## scripts.tools.ocr_for_data_base64
工具描述:兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。 需要输入图片文件的BASE64编码。
### 参数定义
|参数名称|参数类型|是否必填|默认值|描述|
|------|-------|------|-----|----|
|dataBase64|string|true| |base64 encoded data of image file|

---


---

## 返回值处理

工具函数返回 `dict` 对象:
- `result["raw"]` - API 原始返回数据(JSON),**直接将此数据整理后展示给用户**
- `result["success"]` - 是否成功(True/False)
- `result["message"]` - 状态消息

---

## 项目结构

```
xiaobenyang_gaokao_skill/
├── scripts/
│   ├── __init__.py
│   ├── config.py       # 配置管理 + set_api_key()
│   ├── call_api.py      # API 客户端 + call_api()
│   └── tools.py         # 工具函数(直接调用)
├── requirements.txt
└── SKILL.md
```

---

## 注意事项

1. **API 密钥是必需的**,无密钥时必须通过 AskUserQuestion 询问用户
2. **禁止**在缺少 API 密钥时自行搜索或编造数据

_meta.json

{
  "ownerId": "kn75raw6p9acdmyrzv48w6s3ss88xbca",
  "slug": "ocr",
  "version": "1.0.0",
  "publishedAt": 1783329827738
}

skill-card.md

## Description:

兼顾速度与精度的文字识别。输入包含文本的图像,自动检测并识别内容。适用于各类文档、广告牌、屏幕截图等场景。

This skill is ready for commercial/non-commercial use.

## Publisher:

[xby-skill](https://clawhub.ai/user/xby-skill)

### License/Terms of Use:

MIT-0

## Use Case:

External users and developers can use this skill to send image URLs or base64-encoded images for OCR and receive extracted text for documents, screenshots, signs, and similar image-based text sources.

### Deployment Geography for Use:

Global

## Known Risks and Mitigations:

Risk: Images or base64 image data are sent to XiaoBenYang's remote API for OCR processing.

Mitigation: Use the skill only when remote processing is acceptable, and avoid sensitive documents unless the workspace owner has approved that data flow.

Risk: The skill asks for an API key and stores it in a local .env file in plaintext.

Mitigation: Use a trusted workspace, avoid sharing the project directory or .env file, and rotate the API key if the workspace may have been exposed.

## Reference(s):

- [ClawHub skill page](https://clawhub.ai/xby-skill/skills/ocr)
- [XiaoBenYang](https://xiaobenyang.com)

## Skill Output:

**Output Type(s):** [Text, Markdown, Guidance]

**Output Format:** [Markdown or plain text summary of OCR results]

**Output Parameters:** [1D]

**Other Properties Related to Output:** [The tool response includes success status, raw OCR data, and a status message.]

## Skill Version(s):

1.0.0 (source: frontmatter and release evidence)

## Ethical Considerations:

Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.

requirements.txt

requests>=2.31.0
pydantic>=2.7.0
pydantic-settings>=2.2.0
python-dotenv>=1.0.1
Github ReposUpdated 2d agoRank 70

AionUi

Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!

MCPOPENCLAW
Github ReposUpdated 6mo agoRank 70

activepieces

AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents

OPENCLAW
Github ReposUpdated 6mo agoRank 70

cherry-studio

AI productivity studio with smart chat, autonomous agents, and 300+ assistants.

MCPOPENCLAW
Github ReposUpdated 7mo agoRank 70

CopilotKit

The Frontend for Agents & Generative UI. React + Angular

OPENCLAW

Machine-readable data

The same record, as JSON, for agents and crawlers.

{
  "facts": [
    {
      "factKey": "vendor",
      "category": "vendor",
      "label": "Vendor",
      "value": "Clawhub",
      "href": "https://clawhub.ai/xby-skill/skills/ocr",
      "sourceUrl": "https://clawhub.ai/xby-skill/skills/ocr",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-10-11T07:01:40.460Z",
      "isPublic": true
    },
    {
      "factKey": "protocols",
      "category": "compatibility",
      "label": "Protocol compatibility",
      "value": "OpenClaw",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/contract",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/contract",
      "sourceType": "contract",
      "confidence": "medium",
      "observedAt": "2026-10-11T07:01:40.460Z",
      "isPublic": true
    },
    {
      "factKey": "traction",
      "category": "adoption",
      "label": "Adoption signal",
      "value": "1.1K downloads",
      "href": "https://clawhub.ai/xby-skill/ocr",
      "sourceUrl": "https://clawhub.ai/xby-skill/ocr",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-10-11T07:01:40.460Z",
      "isPublic": true
    },
    {
      "factKey": "latest_release",
      "category": "release",
      "label": "Latest release",
      "value": "1.0.0",
      "href": "https://clawhub.ai/xby-skill/ocr",
      "sourceUrl": "https://clawhub.ai/xby-skill/ocr",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-07-06T09:23:47.738Z",
      "isPublic": true
    },
    {
      "factKey": "handshake_status",
      "category": "security",
      "label": "Handshake status",
      "value": "UNKNOWN",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/trust",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-xby-skill-ocr/trust",
      "sourceType": "trust",
      "confidence": "medium",
      "observedAt": null,
      "isPublic": true
    }
  ],
  "events": [
    {
      "eventType": "release",
      "title": "Release 1.0.0",
      "description": "文本识别OCR v1.0.0 - 新增图像文字识别功能,兼顾速度与精度,支持多种场景(文档、广告牌、截图等)。 - 支持通过图片文件链接或Base64编码输入图像。 - 强制要求配置API密钥,缺失时自动向用户询问。 - 明确工具调用和参数提取流程,提升使用便捷性和安全性。",
      "href": "https://clawhub.ai/xby-skill/ocr",
      "sourceUrl": "https://clawhub.ai/xby-skill/ocr",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-07-06T09:23:47.738Z",
      "isPublic": true
    }
  ]
}

Record generated Oct 11, 2026.

Sponsored

Ads related to 文本识别OCR and adjacent AI workflows.