A comprehensive command-line tool for parsing Microsoft Word documents and extracting text, tables, and metadata into structured formats.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install word-reader
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install word-reader using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
The Word Document Reader is a specialized utility designed to bridge the gap between binary document formats and structured data. It enables users to programmatically access the contents of both modern .docx and legacy .doc files, making it an essential component for automated document processing within Openclaw Skills.
By leveraging the python-docx library, this tool accurately identifies document structures such as headers, footers, and nested tables. It provides a flexible way for developers and researchers to convert unstructured document data into clean Markdown, plain text, or machine-readable JSON, facilitating easier integration with AI models and data pipelines.
To get started with this skill for Openclaw Skills, ensure you have Python 3 installed and then follow these steps:
# Install the core Python dependency
pip3 install python-docx
# (Optional) Install antiword for legacy .doc support on Linux
sudo apt-get install antiword -y
# (Optional) Install antiword for legacy .doc support on macOS
brew install antiword
The tool organizes document data into a structured schema, especially when using the JSON output format:
| Key | Type | Description |
|---|---|---|
metadata |
object | Contains document properties like title, author, and timestamps |
text |
string | The complete text content extracted from the document |
tables |
array | An array of tables, where each table is a 2D array of strings |
images |
array | Metadata for embedded images, including filename and dimensions |
format |
string | The output format requested (json, text, or markdown) |
Loading
A comprehensive tool for generating professional AI images and videos using Seedream 4.5 and Kling models via the LiblibAI API.

A powerful integration for Google Gemini that enables end-to-end multimodal workflows including image generation, video synthesis, and speech-to-text capabilities.

An automated tool for extracting and documenting UI/UX design systems from frontend codebases while stripping away business logic.

A professional AI tool for generating high-resolution images via Alibaba's wan2.6-t2i model with automated local storage and cloud hosting.

A high-performance web navigation tool designed to bypass advanced anti-bot systems using stealth-optimized browser instances.

A bridge between AI agents and ComfyUI that allows users to programmatically trigger workflows and retrieve generated image assets.








































