A comprehensive document processing suite for AI agents to manipulate PDFs, perform OCR, and generate text-to-speech audio.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install pdf-toolkit
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install pdf-toolkit using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
The PDF Toolkit is a versatile utility designed to empower AI agents with deep document manipulation capabilities. By leveraging this tool within the Openclaw Skills ecosystem, agents can programmatically interact with PDF and DOCX files, performing tasks ranging from simple metadata extraction to advanced optical character recognition (OCR). It utilizes the uv python manager for seamless execution, ensuring dependencies are handled automatically in isolated environments.
This skill is particularly valuable for developers who need their agents to process unstructured data, automate document workflows, or provide accessibility features like text-to-speech. It operates locally on the host machine, providing high performance and security, while optionally utilizing external services for high-quality audio generation. Integrating this into your Openclaw Skills library transforms a standard AI agent into a sophisticated document processing assistant.
To use this skill, ensure that uv is installed on your system. Certain advanced features require additional host binaries.
# Check which features are currently supported on your host
uv run src/main.py doctor
# Optional: Install system dependencies for OCR and TTS
# macOS
brew install tesseract poppler ffmpeg
# Ubuntu/Debian
sudo apt install tesseract-ocr poppler-utils ffmpeg
The PDF Toolkit manages data through standard file system paths and structured CLI output.
| Feature | Input Type | Output Format |
|---|---|---|
| Info | PDF Path | Plain text (Metadata, Page Count) |
| Extraction | PDF Path | Plain text, CSV (Tables), or Image files |
| OCR | Scanned PDF | Searchable text or PDF |
| TTS | Text or PDF | MP3 Audio file |
| Conversion | Document Path | Target format (PDF, DOCX, etc.) |
Loading
An automated financial analysis workflow that orchestrates price tracking, fundamental data retrieval, and market sentiment into a unified equity research report.

MusiClaw is an AI music production skill that allows agents to generate instrumental beats, list them on a commercial marketplace, and manage earnings autonomously.

A comprehensive webchat interface for OpenClaw that supports interactive UI widgets, 3D voxel avatars, and integrated task management.

A professional toolset for automating PDF and Word document conversions, editing, and OCR processing.

A local, privacy-first compliance tool to detect prohibited and sensitive keywords in short video scripts and livestreams without requiring an API key.

A professional citation management tool for automating authentic bibliographic metadata retrieval and standardizing research paper references across multiple international styles.








































