Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale. Skill: Pdf Owner: awspace Summary: Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale. Tags: latest:0.1.0 Version history: v0.1.0 | 2026-01-31T08:24:15.729Z | auto - Initial release of the PDF manipulation toolkit. - Su
Rank
62
Safety
84
Downloads
36k
Updated
May 16, 2026
Version
0.1.0
Source
CLAWHUB
About
What it does, and when to use it.
Capability contract not published. No trust telemetry is available yet. 36.2K downloads reported by the source. Last updated 5/16/2026.
Avoid when
- Contract metadata is missing or unavailable for deterministic execution.
Risk flags: missing_or_unavailable_contract, trust_data_unavailable, schema_references_missing
Public facts
Every fact links back to the source it came from.
- Vendor
- Clawhubvendor · observed May 16, 2026
- Protocol compatibility
- OpenClawcompatibility · observed May 16, 2026
- Adoption signal
- 36.2K downloadsadoption · observed May 16, 2026
- Latest release
- 0.1.0release · observed Jan 31, 2026
- Handshake status
- UNKNOWNsecurity
Install and run
Setup complexity: low.
clawhub skill install publishers:awspace:pdf- Setup complexity is LOW. This package is likely designed for quick installation with minimal external side-effects.
- Final validation: Expose the agent to a mock request payload inside a sandbox and trace the network egress before allowing access to real customer data.
Contract: missing
curl -s "https://www.xpersona.co/api/v1/agents/clawhub-awspace-pdf-2/snapshot"
Documentation
CLAWHUB
8,215 characters of source documentation, loaded on request.
Extracted files
2 files captured from the source.
SKILL.md
---
name: pdf
description: Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.
license: Proprietary. LICENSE.txt has complete terms
---
# PDF Processing Guide
## Overview
This guide covers essential PDF processing operations using Python libraries and command-line tools. For advanced features, JavaScript libraries, and detailed examples, see reference.md. If you need to fill out a PDF form, read forms.md and follow its instructions.
## Quick Start
```python
from pypdf import PdfReader, PdfWriter
# Read a PDF
reader = PdfReader("document.pdf")
print(f"Pages: {len(reader.pages)}")
# Extract text
text = ""
for page in reader.pages:
text += page.extract_text()
```
## Python Libraries
### pypdf - Basic Operations
#### Merge PDFs
```python
from pypdf import PdfWriter, PdfReader
writer = PdfWriter()
for pdf_file in ["doc1.pdf", "doc2.pdf", "doc3.pdf"]:
reader = PdfReader(pdf_file)
for page in reader.pages:
writer.add_page(page)
with open("merged.pdf", "wb") as output:
writer.write(output)
```
#### Split PDF
```python
reader = PdfReader("input.pdf")
for i, page in enumerate(reader.pages):
writer = PdfWriter()
writer.add_page(page)
with open(f"page_{i+1}.pdf", "wb") as output:
writer.write(output)
```
#### Extract Metadata
```python
reader = PdfReader("document.pdf")
meta = reader.metadata
print(f"Title: {meta.title}")
print(f"Author: {meta.author}")
print(f"Subject: {meta.subject}")
print(f"Creator: {meta.creator}")
```
#### Rotate Pages
```python
reader = PdfReader("input.pdf")
writer = PdfWriter()
page = reader.pages[0]
page.rotate(90) # Rotate 90 degrees clockwise
writer.add_page(page)
with open("rotated.pdf", "wb") as output:
writer.write(output)
```
### pdfplumber - Text and Table Extraction
#### Extract Text with Layout
```python
import pdfplumber
with pdfplumber.open("document.pdf") as pdf:
for page in pdf.pages:
text = page.extract_text()
print(text)
```
#### Extract Tables
```python
with pdfplumber.open("document.pdf") as pdf:
for i, page in enumerate(pdf.pages):
tables = page.extract_tables()
for j, table in enumerate(tables):
print(f"Table {j+1} on page {i+1}:")
for row in table:
print(row)
```
#### Advanced Table Extraction
```python
import pandas as pd
with pdfplumber.open("document.pdf") as pdf:
all_tables = []
for page in pdf.pages:
tables = page.extract_tables()
for table in tables:
if table: # Check if table is not empty
df = pd.DataFrame(table[1:], columns=table[0])
all_tables.append(df)
# Combine all tables
if all_tables:
combined_df = pd.concat(all_tables, ignore_index=True)
combined_df.to_excel("extracted_table_meta.json
{
"ownerId": "kn725kyxpr1g1e3rjqtkp26v0s809930",
"slug": "pdf",
"version": "0.1.0",
"publishedAt": 1769847855729
}AionUi
Free, local, open-source 24/7 Cowork app and OpenClaw for Gemini CLI, Claude Code, Codex, OpenCode, Qwen Code, Goose CLI, Auggie, and more | 🌟 Star if you like it!
activepieces
AI Agents & MCPs & AI Workflow Automation • (~400 MCP servers for AI agents) • AI Automation / AI Agent with MCPs • AI Workflows & AI Agents • MCPs for AI Agents
cherry-studio
AI productivity studio with smart chat, autonomous agents, and 300+ assistants.
CopilotKit
The Frontend for Agents & Generative UI. React + Angular
Machine-readable data
The same record, as JSON, for agents and crawlers.
{
"facts": [
{
"factKey": "vendor",
"label": "Vendor",
"value": "Clawhub",
"category": "vendor",
"href": "https://clawhub.ai/awspace/pdf",
"sourceUrl": "https://clawhub.ai/awspace/pdf",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-05-16T06:12:13.215Z",
"isPublic": true,
"metadata": {}
},
{
"factKey": "protocols",
"label": "Protocol compatibility",
"value": "OpenClaw",
"category": "compatibility",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-awspace-pdf-2/contract",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-awspace-pdf-2/contract",
"sourceType": "contract",
"confidence": "medium",
"observedAt": "2026-05-16T06:12:13.215Z",
"isPublic": true,
"metadata": {}
},
{
"factKey": "traction",
"label": "Adoption signal",
"value": "36.2K downloads",
"category": "adoption",
"href": "https://clawhub.ai/awspace/pdf",
"sourceUrl": "https://clawhub.ai/awspace/pdf",
"sourceType": "profile",
"confidence": "medium",
"observedAt": "2026-05-16T06:12:13.215Z",
"isPublic": true,
"metadata": {}
},
{
"factKey": "latest_release",
"label": "Latest release",
"value": "0.1.0",
"category": "release",
"href": "https://clawhub.ai/awspace/pdf",
"sourceUrl": "https://clawhub.ai/awspace/pdf",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-01-31T08:24:15.729Z",
"isPublic": true,
"metadata": {}
},
{
"factKey": "handshake_status",
"label": "Handshake status",
"value": "UNKNOWN",
"category": "security",
"href": "https://www.xpersona.co/api/v1/agents/clawhub-awspace-pdf-2/trust",
"sourceUrl": "https://www.xpersona.co/api/v1/agents/clawhub-awspace-pdf-2/trust",
"sourceType": "trust",
"confidence": "medium",
"observedAt": null,
"isPublic": true,
"metadata": {}
}
],
"events": [
{
"eventType": "release",
"title": "Release 0.1.0",
"description": "- Initial release of the PDF manipulation toolkit. - Supports text and table extraction, PDF creation, merging, splitting, metadata extraction, and rotation. - Guides for handling scanned PDFs (OCR), password protection, watermarking, and image extraction. - Provides command-line tool usage (pdftotext, qpdf, pdftk, pdfimages). - Recipes and code examples for Python libraries: pypdf, pdfplumber, and reportlab. - Quick reference and next steps included for users needing advanced or form-related workflows.",
"href": "https://clawhub.ai/awspace/pdf",
"sourceUrl": "https://clawhub.ai/awspace/pdf",
"sourceType": "release",
"confidence": "medium",
"observedAt": "2026-01-31T08:24:15.729Z",
"isPublic": true,
"metadata": {}
}
]
}Record generated Oct 9, 2026.
