PDF Translator for Openclaw

A sophisticated translation engine that preserves complex document layouts by leveraging LaTeX reconstruction and direct arXiv source integration.

overdue-lin
v1.0.0
Apr 4, 2026
0
588
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install pdf-translate-skill

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install pdf-translate-skill using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is PDF Translator?

The PDF Translator is a high-performance Openclaw Skills utility designed to bridge the gap between static documents and multilingual accessibility. Unlike standard OCR tools, this skill understands the structural integrity of a document, utilizing two distinct workflows to ensure that translated outputs maintain the original visual context, column structures, and embedded imagery.

By combining multimodal vision analysis with professional LaTeX typesetting, it serves as an end-to-end pipeline for researchers and professionals. Whether processing a local file or pulling directly from the arXiv repository, the skill ensures that technical terms, mathematical formulas, and bibliographies remain intact and professionally formatted.

PDF Translator Use Cases

  • Translating academic arXiv papers into Chinese or other languages while preserving original TeX formatting.
  • Converting local business reports and technical manuals into translated PDFs with identical layout structures.
  • Automating the extraction and translation of embedded text in image-heavy PDF documents.
  • Reconstructing complex multi-column document layouts into editable LaTeX formats.

How PDF Translator Works

  1. Input Detection: The skill identifies if the source is a local PDF file or an arXiv ID/URL.
  2. Source Acquisition: For arXiv, it downloads the original TeX source; for local files, it converts pages into high-resolution images for analysis.
  3. Content Extraction: It extracts text, images, and layout metadata, identifying non-translatable elements like math formulas and citations.
  4. Multimodal Translation: The AI translates text blocks while preserving LaTeX markup or structural hierarchy.
  5. LaTeX Reconstruction: It generates or modifies LaTeX code to wrap the translated content in a format that mirrors the original document.
  6. Multi-Pass Compilation: The system runs XeLaTeX three times to resolve all cross-references, citations, and bibliographies into a final PDF.

PDF Translator Setup

To utilize this skill within the Openclaw Skills ecosystem, ensure the following dependencies are installed:

# Install Python dependencies
pip install pymupdf pillow requests

# Install TeX Live (Debian/Ubuntu)
sudo apt install texlive-xetex texlive-lang-chinese

# Install TeX Live (macOS)
brew install mactex

PDF Translator Data Schema & Taxonomy

The skill manages data through a structured working directory to ensure traceability and easy debugging:

Component Description
pages/ High-quality PNG renders of each PDF page (150 DPI).
extracted/ Individual image assets extracted from the original PDF.
manifest.json Metadata containing page dimensions and layout coordinates.
translated.tex The final generated LaTeX source file.
output.pdf The compiled, translated document.

PDF Translator Advanced Features

  • Direct arXiv Integration: Automatically fetches and unpacks .tar.gz source files for pixel-perfect academic translation.
  • Multi-Pass Resolution: Built-in triple-pass compilation logic to ensure \cite and \ref tags are never broken.
  • Multimodal Layout Analysis: Uses vision-based AI to detect multi-column shifts and complex header/footer patterns.
  • Chinese Language Optimization: Native support for the ctex package and XeLaTeX engine for high-quality Asian character rendering.
  • Intelligent Fallback: Automatically switches to image-based layout reconstruction if TeX source files are unavailable.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*