A high-performance OCR skill for extracting text from images, screenshots, and scanned documents in over 100 languages using PaddleOCR.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install smar
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install smar using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
Smart OCR is a robust text-extraction tool designed for the Openclaw Skills framework. It leverages the state-of-the-art PaddleOCR engine to identify and convert visual text into machine-readable data. Whether you are dealing with complex business cards, handwritten notes, or multi-page scanned PDFs, this skill provides the precision needed for modern AI workflows.
By integrating this tool into your Openclaw Skills setup, you gain access to features like angle classification, layout reconstruction, and multilingual support. It is optimized for both CPU and GPU environments, ensuring fast processing times for single images or large batches of documents.
To get started with this skill in your Openclaw Skills environment, install the necessary dependencies:
# Install the OCR engine (CPU version)
pip install paddlepaddle paddleocr
# For GPU acceleration (optional)
pip install paddlepaddle-gpu
# Install additional document processing tools
pip install pdf2image Pillow
The Smart OCR skill organizes its output into structured data, tracking spatial coordinates and text accuracy.
| Attribute | Description |
|---|---|
| text | The actual string extracted from the image. |
| confidence | A decimal score (0.0 - 1.0) representing extraction accuracy. |
| bbox | Bounding box coordinates (left, top, right, bottom). |
| raw_box | The four-point polygon vertices of the text area. |
| language | The detected or specified language code used for OCR. |
Loading
A high-performance skill for generating AI images using industry-leading models like FLUX and Kolors through the SiliconFlow API.

A specialized tool for listing, finding, and searching Intercom customer conversations via a structured JSON interface.

A powerful tool for discovering and executing Apify Actors to scrape structured data from social media and websites.

A high-performance local image generation skill that interfaces with DrawThings to provide Stable Diffusion capabilities via an API on macOS.

An AI-powered cinema assistant that bridges TMDB data with Emby and Plex servers for intelligent media management and context-aware recommendations.

A dedicated slow-channel notification system that queues non-urgent agent updates for structured human review.








































