Playwright Scraper Skill for Openclaw

A high-performance web scraping tool for Openclaw Skills designed to bypass advanced anti-bot protections like Cloudflare using Playwright Stealth.

waisimon
v1.2.0
Feb 7, 2026
57
31.5k
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install playwright-scraper-skill

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install playwright-scraper-skill using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is Playwright Scraper Skill?

The Playwright Scraper Skill is a versatile data extraction engine built for the Openclaw Skills ecosystem. It provides developers with a tiered approach to web scraping, allowing them to choose the most efficient method based on a target website's security level. From simple static fetches to bypassing sophisticated Cloudflare challenges, this skill ensures that your AI agents can access the web data they need without being blocked.

By leveraging the power of Playwright, this skill mimics human browser behavior more effectively than traditional scraping libraries. It includes specialized scripts for stealth operations, making it an essential addition to any Openclaw Skills collection for researchers, developers, and data analysts who require reliable access to dynamic or protected web content.

Playwright Scraper Skill Use Cases

  • Extracting content from JavaScript-heavy dynamic websites that require full browser rendering.
  • Bypassing high-level anti-bot protections and Cloudflare challenges on community forums and news sites.
  • Automating the collection of structured data for competitive analysis and research within Openclaw Skills workflows.
  • Generating visual proof of page states via automated screenshots and HTML archival.

How Playwright Scraper Skill Works

  1. The user identifies the target URL and assesses the anti-bot level of the destination site.
  2. The skill selects the appropriate execution path: built-in fetch for simple sites, Playwright Simple for JS-heavy sites, or Playwright Stealth for protected sites.
  3. In stealth mode, the skill initializes a browser instance that hides automation signatures like the navigator.webdriver property.
  4. Realistic User-Agents and human-like interaction patterns, such as random delays, are applied to the session.
  5. The browser navigates to the target, waits for the DOM or network to become idle, and extracts the page content.
  6. The processed data is returned to the Openclaw Skills environment as a structured JSON object, including metadata and optional file paths for screenshots.

Playwright Scraper Skill Setup

To integrate this capability into your Openclaw Skills setup, follow these installation steps:

cd playwright-scraper-skill
npm install
npx playwright install chromium

Playwright Scraper Skill Data Schema & Taxonomy

The Playwright Scraper Skill organizes its output into a structured format to ensure compatibility with other Openclaw Skills.

Attribute Type Description
url String The target URL that was scraped.
title String The page title extracted from the metadata.
content String The primary text or HTML content of the page.
elapsedSeconds String Total time taken to execute the scrape.
screenshot Path (Optional) Local path to the generated PNG screenshot.
html_file Path (Optional) Local path to the saved raw HTML source.

Playwright Scraper Skill Advanced Features

  • Stealth execution using addInitScript to mask framework signatures before page load.
  • Environment variable support for customizing wait times, User-Agents, and headless/headful modes.
  • Success-verified templates for challenging domains like Discuss.com.hk, achieving 100% bypass rates.
  • Modular architecture that allows for easy integration with specialized Openclaw Skills for platforms like YouTube and Reddit.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*