Ocr Document
OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten n... Skill: Ocr Document Owner: tanis90 Summary: OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten n... Tags: latest:1.0.0 Version history: v1.0.0 | 2026-03-24T13:53:06.179Z | auto - Initial release of ocr_document skill for OCR text extraction from scanned documents, images, and handwritten notes. - Supports PDF, PNG
Rank
62
Safety
84
Downloads
4.3k
Updated
Oct 9, 2026
Version
1.0.0
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 4.3K downloads reported by the source. Last updated 10/9/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed Oct 9, 2026
- Protocol compatibility
- OpenClawcompatibility · observed Oct 9, 2026
- Adoption signal
- 4.3K downloadsadoption · observed Oct 9, 2026
- Latest release
- 1.0.0release · observed Mar 24, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install s17c6n53ezjj2w372fjv34m7zx83gva1:ocr-document- Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/snapshot"
Documentation
CLAWHUB
5,605 characters of source documentation, loaded on request.
Extracted files
3 files captured from the source.
SKILL.md
---
name: ocr_document
description: "OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten notes, or any document that needs optical character recognition."
homepage: https://mineru.net
metadata: {"openclaw":{"emoji":"🔍","requires":{"bins":["mineru-open-api"]},"install":[{"id":"npm","kind":"node","package":"mineru-open-api","bins":["mineru-open-api"],"label":"Install via npm"},{"id":"uv","kind":"uv","package":"mineru-open-api","bins":["mineru-open-api"],"label":"Install via uv"},{"id":"go","kind":"go","package":"github.com/opendatalab/MinerU-Ecosystem/cli/mineru-open-api","bins":["mineru-open-api"],"label":"Install via go install","os":["darwin","linux"]}]}}
---
# OCR Document - Extract Text from Scanned Documents and Images
Extract text from scanned documents and images using OCR via MinerU Open API. No API key required.
## Quick Start
```bash
# OCR a scanned PDF
mineru-open-api flash-extract scanned.pdf
# OCR an image of a document
mineru-open-api flash-extract page-photo.jpg
# OCR from URL (no download needed)
mineru-open-api flash-extract https://example.com/scanned.pdf
# Specify language for better accuracy
mineru-open-api flash-extract scanned.pdf --language en
# Save OCR result to file
mineru-open-api flash-extract scanned.pdf -o ./output/
```
## Language Rule
You MUST reply to the user in the SAME language they use. This is non-negotiable.
## Capabilities
- OCR for scanned PDFs, photographed documents, images
- Supports PDF, PNG, JPG, WebP, BMP, TIFF
- Supports both local files and URLs directly
- Language hint with `--language` (default: `ch`, use `en` for English)
- No API key, no signup, no authentication
- Max 10MB / 20 pages per document
## When to Use
- User asks to "OCR" a document or image
- User has a scanned PDF that needs text extraction
- User shares a photo of a page and wants the text
- User mentions "scan", "handwriting", or "recognize text"
## CLI Reference
Run `mineru-open-api flash-extract --help` for all available options.
## Data Privacy
- `flash-extract` uploads the document to MinerU's cloud API for processing and returns the result. No account or API key is required.
- Documents are processed in real-time and are not stored after extraction.
- For details, see https://mineru.net
## Notes
- Best results with clear, high-resolution scans
- For higher precision OCR with full layout preservation, use `mineru-open-api extract --ocr` (requires auth via `mineru-open-api auth`)
- If the CLI cannot be installed via npm/uv/go, download it from https://mineru.net/ecosystem?tab=cli_meta.json
{
"ownerId": "kn7ftqm5yamemf6n2473s634v183g5rq",
"slug": "ocr-document",
"version": "1.0.0",
"publishedAt": 1774360386179
}skill-card.md
## Description: OCR document extraction - extract text from scanned documents, photos, and images using OCR. This skill is ready for commercial/non-commercial use. ## Publisher: [tanis90](https://clawhub.ai/user/tanis90) ### License/Terms of Use: MIT-0 ## Use Case: Developers and external users use this skill to extract text from scanned PDFs, photographed pages, images, and handwritten notes through the MinerU Open API CLI. ### Deployment Geography for Use: Global ## Known Risks and Mitigations: Risk: The skill installs and uses a third-party CLI. Mitigation: Install only in environments where the MinerU CLI is approved, and prefer a reviewed or pinned CLI version in controlled deployments. Risk: Documents are uploaded to MinerU's cloud OCR service for processing. Mitigation: Avoid sending confidential documents unless MinerU's privacy terms meet the user's requirements. ## Reference(s): - [ClawHub Skill Page](https://clawhub.ai/tanis90/skills/ocr-document) - [MinerU](https://mineru.net) - [MinerU CLI Download](https://mineru.net/ecosystem?tab=cli) - [mineru-open-api npm package](https://www.npmjs.com/package/mineru-open-api) ## Skill Output: **Output Type(s):** [Text, Markdown, Shell commands, Guidance] **Output Format:** [Markdown guidance with inline shell commands and OCR text output] **Output Parameters:** [1D] **Other Properties Related to Output:** [Supports local files and URLs; documented limits are 10MB or 20 pages per document.] ## Skill Version(s): 1.0.0 (source: server release evidence) ## Ethical Considerations: Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"category": "vendor",
"label": "Vendor",
"value": "Clawhub",
"href": "https://clawhub.ai/tanis90/skills/ocr-document",
"sourceUrl": "https://clawhub.ai/tanis90/skills/ocr-document",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-09T05:39:38.708Z",
"isPublic": true
},
{
"factKey": "protocols",
"category": "compatibility",
"label": "Protocol compatibility",
"value": "OpenClaw",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-10-09T05:39:38.708Z",
"isPublic": true
},
{
"factKey": "traction",
"category": "adoption",
"label": "Adoption signal",
"value": "4.3K downloads",
"href": "https://clawhub.ai/tanis90/ocr-document",
"sourceUrl": "https://clawhub.ai/tanis90/ocr-document",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-10-09T05:39:38.708Z",
"isPublic": true
},
{
"factKey": "latest_release",
"category": "release",
"label": "Latest release",
"value": "1.0.0",
"href": "https://clawhub.ai/tanis90/ocr-document",
"sourceUrl": "https://clawhub.ai/tanis90/ocr-document",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-03-24T13:53:06.179Z",
"isPublic": true
},
{
"factKey": "handshake_status",
"category": "security",
"label": "Handshake status",
"value": "UNKNOWN",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true
}
],
"events": [
{
"eventType": "release",
"title": "Release 1.0.0",
"description": "- Initial release of ocr_document skill for OCR text extraction from scanned documents, images, and handwritten notes. - Supports PDF, PNG, JPG, WebP, BMP, and TIFF formats from local files or URLs. - No API key, signup, or authentication required. - Language selection available for improved accuracy; replies always match the user's language. - Maximum file size is 10MB or 20 pages per document. - Powered by the MinerU Open API CLI; installation guides provided for npm, uv, go, and direct download.",
"href": "https://clawhub.ai/tanis90/ocr-document",
"sourceUrl": "https://clawhub.ai/tanis90/ocr-document",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-03-24T13:53:06.179Z",
"isPublic": true
}
]
}Record generated Oct 9, 2026.
