Diffbot Fetch for Openclaw

A high-performance extraction tool that converts web articles into clean, readable Markdown using the Diffbot Article API.

flobo3
v1.0.0
Mar 31, 2026
0
607
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install diffbot-fetch

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install diffbot-fetch using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is Diffbot Fetch?

Diffbot Fetch is a specialized utility within the Openclaw Skills ecosystem designed to bypass web clutter and extract the signal from the noise. By leveraging the Diffbot Article API, this skill identifies the primary content of any webpage—be it a blog post, news report, or long-form essay—and strips away intrusive ads, navigation menus, and sidebars. The result is a structured, clean Markdown output that is perfectly formatted for AI consumption or archival purposes.

Diffbot Fetch Use Cases

  • Extracting clean text from news articles for automated summarization.
  • Scraping blog posts to build local knowledge bases or RAG systems.
  • Converting cluttered web pages into readable Markdown for offline consumption.
  • Gathering research data without the overhead of HTML boilerplate.

How Diffbot Fetch Works

  1. The user provides a target URL to the Diffbot Fetch skill via the CLI or agent interface.
  2. The skill sends a request to the Diffbot Article API using the configured API token.
  3. Diffbot's computer vision and NLP models analyze the page to identify the core article body.
  4. The raw extraction is processed and converted into a clean Markdown format.
  5. The finalized content is returned to the user or the calling Openclaw Skills agent for further processing.

Diffbot Fetch Setup

To get started with this entry in the Openclaw Skills library, ensure you have a valid Diffbot API token. Configure your environment as follows:

# Set your Diffbot API Key
export DIFFBOT_API_KEY="your_token_here"

# Run the fetch command
uv run fetch.py "https://example.com/article"

Diffbot Fetch Data Schema & Taxonomy

The skill focuses on a streamlined output schema to ensure compatibility with other Openclaw Skills.

Field Type Description
Source URL String The original link provided for extraction.
Content Markdown The primary article body stripped of HTML tags.
Metadata Object Includes available author info, date, and title where applicable.

Diffbot Fetch Advanced Features

  • Intelligent identification of multi-page articles for comprehensive extraction.
  • Automatic conversion of HTML tables and lists into standard Markdown syntax.
  • Seamless integration with other Openclaw Skills for automated content pipelines.
  • High-accuracy extraction powered by Diffbot's proprietary machine learning models.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*