Claim this agent
agentCLAWHUBUnverified

Ocr Document

OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten n... Skill: Ocr Document Owner: tanis90 Summary: OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten n... Tags: latest:1.0.0 Version history: v1.0.0 | 2026-03-24T13:53:06.179Z | auto - Initial release of ocr_document skill for OCR text extraction from scanned documents, images, and handwritten notes. - Supports PDF, PNG

OpenClaw

Rank

62

Safety

84

Downloads

4.3k

Updated

Oct 9, 2026

Version

1.0.0

Source

CLAWHUB

About

What it does, and when to use it.

Capability contract not published. No trust telemetry is available yet. 4.3K downloads reported by the source. Last updated 10/9/2026.

Avoid when

  • Contract metadata is missing or unavailable for deterministic execution.

Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing

Public facts

Every fact links back to the source it came from.

Vendor
Clawhubvendor · observed Oct 9, 2026
Protocol compatibility
OpenClawcompatibility · observed Oct 9, 2026
Adoption signal
4.3K downloadsadoption · observed Oct 9, 2026
Latest release
1.0.0release · observed Mar 24, 2026
Handshake status
UNKNOWNsecurity

Install and run

Setup complexity: low.

clawhub skill install s17c6n53ezjj2w372fjv34m7zx83gva1:ocr-document
  1. Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
  2. Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.

Contract: missing

curl -s "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/snapshot"

Documentation

CLAWHUB

5,605 characters of source documentation, loaded on request.

Extracted files

3 files captured from the source.

SKILL.md

---
name: ocr_document
description: "OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten notes, or any document that needs optical character recognition."
homepage: https://mineru.net
metadata: {"openclaw":{"emoji":"🔍","requires":{"bins":["mineru-open-api"]},"install":[{"id":"npm","kind":"node","package":"mineru-open-api","bins":["mineru-open-api"],"label":"Install via npm"},{"id":"uv","kind":"uv","package":"mineru-open-api","bins":["mineru-open-api"],"label":"Install via uv"},{"id":"go","kind":"go","package":"github.com/opendatalab/MinerU-Ecosystem/cli/mineru-open-api","bins":["mineru-open-api"],"label":"Install via go install","os":["darwin","linux"]}]}}
---

# OCR Document - Extract Text from Scanned Documents and Images

Extract text from scanned documents and images using OCR via MinerU Open API. No API key required.

## Quick Start

```bash
# OCR a scanned PDF
mineru-open-api flash-extract scanned.pdf

# OCR an image of a document
mineru-open-api flash-extract page-photo.jpg

# OCR from URL (no download needed)
mineru-open-api flash-extract https://example.com/scanned.pdf

# Specify language for better accuracy
mineru-open-api flash-extract scanned.pdf --language en

# Save OCR result to file
mineru-open-api flash-extract scanned.pdf -o ./output/
```

## Language Rule

You MUST reply to the user in the SAME language they use. This is non-negotiable.

## Capabilities

- OCR for scanned PDFs, photographed documents, images
- Supports PDF, PNG, JPG, WebP, BMP, TIFF
- Supports both local files and URLs directly
- Language hint with `--language` (default: `ch`, use `en` for English)
- No API key, no signup, no authentication
- Max 10MB / 20 pages per document

## When to Use

- User asks to "OCR" a document or image
- User has a scanned PDF that needs text extraction
- User shares a photo of a page and wants the text
- User mentions "scan", "handwriting", or "recognize text"

## CLI Reference

Run `mineru-open-api flash-extract --help` for all available options.

## Data Privacy

- `flash-extract` uploads the document to MinerU's cloud API for processing and returns the result. No account or API key is required.
- Documents are processed in real-time and are not stored after extraction.
- For details, see https://mineru.net

## Notes

- Best results with clear, high-resolution scans
- For higher precision OCR with full layout preservation, use `mineru-open-api extract --ocr` (requires auth via `mineru-open-api auth`)
- If the CLI cannot be installed via npm/uv/go, download it from https://mineru.net/ecosystem?tab=cli

_meta.json

{
  "ownerId": "kn7ftqm5yamemf6n2473s634v183g5rq",
  "slug": "ocr-document",
  "version": "1.0.0",
  "publishedAt": 1774360386179
}

skill-card.md

## Description:

OCR document extraction - extract text from scanned documents, photos, and images using OCR.

This skill is ready for commercial/non-commercial use.

## Publisher:

[tanis90](https://clawhub.ai/user/tanis90)

### License/Terms of Use:

MIT-0

## Use Case:

Developers and external users use this skill to extract text from scanned PDFs, photographed pages, images, and handwritten notes through the MinerU Open API CLI.

### Deployment Geography for Use:

Global

## Known Risks and Mitigations:

Risk: The skill installs and uses a third-party CLI.

Mitigation: Install only in environments where the MinerU CLI is approved, and prefer a reviewed or pinned CLI version in controlled deployments.

Risk: Documents are uploaded to MinerU's cloud OCR service for processing.

Mitigation: Avoid sending confidential documents unless MinerU's privacy terms meet the user's requirements.

## Reference(s):

- [ClawHub Skill Page](https://clawhub.ai/tanis90/skills/ocr-document)
- [MinerU](https://mineru.net)
- [MinerU CLI Download](https://mineru.net/ecosystem?tab=cli)
- [mineru-open-api npm package](https://www.npmjs.com/package/mineru-open-api)

## Skill Output:

**Output Type(s):** [Text, Markdown, Shell commands, Guidance]

**Output Format:** [Markdown guidance with inline shell commands and OCR text output]

**Output Parameters:** [1D]

**Other Properties Related to Output:** [Supports local files and URLs; documented limits are 10MB or 20 pages per document.]

## Skill Version(s):

1.0.0 (source: server release evidence)

## Ethical Considerations:

Users should evaluate whether this skill is appropriate for their environment, review any generated or modified files before relying on them, and apply their organization's safety, security, and compliance requirements before deployment.
Github ReposUpdated 1h agoRank 70

AionUi

Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!

MCPOPENCLAW
Github ReposUpdated 6mo agoRank 70

activepieces

AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents

OPENCLAW
Github ReposUpdated 6mo agoRank 70

cherry-studio

AI productivity studio with smart chat, autonomous agents, and 300+ assistants.

MCPOPENCLAW
Github ReposUpdated 7mo agoRank 70

CopilotKit

The Frontend for Agents & Generative UI. React + Angular

OPENCLAW

Machine-readable data

The same record, as JSON, for agents and crawlers.

{
  "facts": [
    {
      "factKey": "vendor",
      "category": "vendor",
      "label": "Vendor",
      "value": "Clawhub",
      "href": "https://clawhub.ai/tanis90/skills/ocr-document",
      "sourceUrl": "https://clawhub.ai/tanis90/skills/ocr-document",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-10-09T05:39:38.708Z",
      "isPublic": true
    },
    {
      "factKey": "protocols",
      "category": "compatibility",
      "label": "Protocol compatibility",
      "value": "OpenClaw",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/contract",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/contract",
      "sourceType": "contract",
      "confidence": "medium",
      "observedAt": "2026-10-09T05:39:38.708Z",
      "isPublic": true
    },
    {
      "factKey": "traction",
      "category": "adoption",
      "label": "Adoption signal",
      "value": "4.3K downloads",
      "href": "https://clawhub.ai/tanis90/ocr-document",
      "sourceUrl": "https://clawhub.ai/tanis90/ocr-document",
      "sourceType": "profile",
      "confidence": "medium",
      "observedAt": "2026-10-09T05:39:38.708Z",
      "isPublic": true
    },
    {
      "factKey": "latest_release",
      "category": "release",
      "label": "Latest release",
      "value": "1.0.0",
      "href": "https://clawhub.ai/tanis90/ocr-document",
      "sourceUrl": "https://clawhub.ai/tanis90/ocr-document",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-03-24T13:53:06.179Z",
      "isPublic": true
    },
    {
      "factKey": "handshake_status",
      "category": "security",
      "label": "Handshake status",
      "value": "UNKNOWN",
      "href": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/trust",
      "sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-tanis90-ocr-document/trust",
      "sourceType": "trust",
      "confidence": "medium",
      "observedAt": null,
      "isPublic": true
    }
  ],
  "events": [
    {
      "eventType": "release",
      "title": "Release 1.0.0",
      "description": "- Initial release of ocr_document skill for OCR text extraction from scanned documents, images, and handwritten notes. - Supports PDF, PNG, JPG, WebP, BMP, and TIFF formats from local files or URLs. - No API key, signup, or authentication required. - Language selection available for improved accuracy; replies always match the user's language. - Maximum file size is 10MB or 20 pages per document. - Powered by the MinerU Open API CLI; installation guides provided for npm, uv, go, and direct download.",
      "href": "https://clawhub.ai/tanis90/ocr-document",
      "sourceUrl": "https://clawhub.ai/tanis90/ocr-document",
      "sourceType": "release",
      "confidence": "medium",
      "observedAt": "2026-03-24T13:53:06.179Z",
      "isPublic": true
    }
  ]
}

Record generated Oct 9, 2026.

Sponsored

Ads related to Ocr Document and adjacent AI workflows.