mineru-precision-extract for Openclaw

A high-precision document extraction tool that converts complex PDFs, images, and office files into structured Markdown, LaTeX, or JSON with advanced OCR and table recognition.

mineru-extract
v0.2.1
Apr 7, 2026
0
1.1k
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install mineru-precision-extract

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install mineru-precision-extract using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is mineru-precision-extract?

mineru-precision-extract is a professional-grade Openclaw Skills integration designed for deep document parsing. It leverages state-of-the-art models to transform static files—including academic papers, scanned PDFs, and multi-page reports—into editable, machine-readable formats. Whether you are building a RAG pipeline or digitizing archives, this skill provides the accuracy needed for production-level data engineering.

The tool offers a unique choice between the vlm model for maximum accuracy in complex layouts and the pipeline model for zero-hallucination reliability. By integrating directly with Openclaw Skills, developers can automate the conversion of 80+ languages and handle massive batch processing tasks through a unified CLI or agent interface.

mineru-precision-extract Use Cases

  • Extracting complex tables and mathematical formulas from academic PDF papers into LaTeX.
  • Converting scanned legal or historical documents into searchable Markdown using high-precision OCR.
  • Automating batch document processing for large datasets in research and data engineering workflows.
  • Crawling web pages to extract structured content while maintaining original layout integrity.
  • Migrating legacy Word (DOCX) or PowerPoint (PPTX) files into structured documentation formats.

How mineru-precision-extract Works

  1. The user provides a local file path, a list of files, or a remote URL to the mineru-open-api.
  2. The skill identifies the document type and applies the selected model (vlm or pipeline) for layout analysis.
  3. OCR and specialized recognition engines extract text, recognize tables, and convert formulas into LaTeX strings.
  4. Extracted data is structured into the requested output formats such as Markdown, HTML, or DOCX.
  5. Images and assets are saved alongside the primary document, and the final result is delivered to the specified output directory.

mineru-precision-extract Setup

To get started with this component of Openclaw Skills, install the CLI via npm or Go:

npm install -g mineru-open-api

Or for macOS/Linux users:

go install github.com/opendatalab/MinerU-Ecosystem/cli/mineru-open-api@latest

After installation, authenticate by creating a token at the MinerU website and running:

mineru-open-api auth

mineru-precision-extract Data Schema & Taxonomy

The skill organizes extracted content into a structured hierarchy. When an output directory is specified, it creates the following taxonomy:

File Type Description
*.md / *.json The primary structured text output containing the document content.
images/ A subdirectory containing all figures and images extracted from the source.
*.html / *.docx Optional alternative formats generated during the extraction process.
metadata Internal taxonomy capturing page counts, language detection, and model logs.

mineru-precision-extract Advanced Features

  • Dual-Model Strategy: Choose vlm for intricate layouts or pipeline for verified, hallucination-free extraction.
  • Massive Batch Processing: Support for processing hundreds of files concurrently via file lists or directory patterns.
  • Intelligent Web Crawling: Convert live URLs directly into clean, structured Markdown for LLM training or archival.
  • Global Language Support: Native recognition for over 80 languages across Latin, CJK, Arabic, and Devanagari scripts.
  • Flexible Page Targeting: Use the --pages flag to extract specific sections from massive documents, saving time and API credits.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Requires
Bins mineru-open-api
Github Stars: 0
forks: 0

Featured*