A robust OCR utility that extracts Chinese and English text from images using the Tesseract engine for AI-driven workflows.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install image-ocr-reader
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install image-ocr-reader using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
The Image OCR Reader is a specialized tool designed to convert visual information into machine-readable text. By leveraging the industry-standard Tesseract OCR engine, it enables developers and AI agents to process image files including JPG, PNG, and JPEG formats with high accuracy. This skill is particularly effective for multi-lingual environments, offering seamless recognition of both Chinese and English characters.
As part of the Openclaw Skills ecosystem, this component acts as a bridge between unstructured image data and structured text processing. Whether you are digitizing archives or building a pipeline for automated document analysis, this skill provides the reliability and speed required for modern AI applications.
# Ubuntu/Debian
sudo apt-get install tesseract-ocr
# macOS
brew install tesseract
# CentOS/RHEL
sudo yum install tesseract
pip install pytesseract Pillow
The Image OCR Reader processes image files and outputs raw text data. Below is the input and output structure:
| Attribute | Description |
|---|---|
| Supported Formats | .jpg, .jpeg, .png |
| Primary Output | UTF-8 encoded string of recognized text |
| Language Support | English (eng), Chinese Simplified (chi_sim) |
| Dependencies | Python 3.x, Tesseract-OCR Engine |
Loading
Openclaw Skills for tRPC provide expert guidance for building end-to-end typesafe APIs without the need for schemas or code generation.

A technical knowledge skill focused on architecting secure, modern, and compliant Stripe payment integrations using the latest API standards.

A versatile command-line interface for managing Google Workspace services like Gmail, Sheets, and Drive via OAuth.

A powerful CLI skill for running, serving, and benchmarking local AI models using containerized runtimes like Docker and Podman.

A lightweight skill to bootstrap a single-target SwiftUI iOS application using XcodeGen for version-control-friendly project management.

A command-line interface skill for controlling Amazon Echo devices and smart home integrations using shell scripts.








































