Browser Execution for Openclaw

Browser Execution completes direct browser tasks with Playwright, computer controls, secure vault handling, recovery tactics, and an explicit verified result.

ojusave
v0.1.0
Sep 17, 2026
0
130
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install browser-execution

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install browser-execution using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is Browser Execution?

Browser Execution is an Openclaw Skills workflow for completing browser-based tasks reliably through Kernel browser sessions. It combines Playwright for navigation, inspection, extraction, and deterministic interaction with computer controls for visual reasoning and coordinate-level input.

The skill emphasizes bounded execution, secure secret handling, human approval for transactions, and clear terminal reporting. It reuses one browser for the full job, responds to blocked pages with limited tactic changes, and calls complete_task exactly once with a verified success or failure outcome.

Browser Execution Use Cases

  • Complete form submissions using user-provided names, contact details, addresses, and other non-credential values.
  • Navigate websites, inspect page content, extract information, and interact with deterministic controls.
  • Use screenshots and computer actions when visual reasoning or coordinate-level interaction is more reliable.
  • Fill saved login, payment, address, phone, identity, or token data without exposing raw secrets to browser calls.
  • Handle blocked pages, timeouts, authentication challenges, and other browser execution failures with bounded recovery.
  • Complete approved purchases or other transactions while validating the merchant, item, quantity, options, fees, and total.
  • Provide an explicit success or failure result instead of leaving a task in an open-ended state.

How Browser Execution Works

  1. Read the current Kernel documentation for browser sessions, Playwright execution, computer controls, and live view rather than relying on assumed API shapes.
  2. Create one browser with manage_browsers and reuse it throughout the task.
  3. Prefer execute_playwright_code for navigation, DOM inspection, extraction, and deterministic interaction; use computer_action when visual or coordinate-based control is more dependable.
  4. Navigate with domcontentloaded or wait for a specific locator, URL, response, or visible state, keeping waits bounded and avoiding fixed sleeps or networkidle.
  5. Use directly supplied non-secret form values as provided. For saved secrets or saved identity data, inspect fields first, obtain an opaque handle with list_vault, and inject values with fill_from_vault.
  6. If a page is blocked, try no more than two materially different relevant approaches, such as a direct provider URL, a different interaction method, or a fresh tab.
  7. For transactions, request approval once for the exact payload and quoted total before filling payment secrets, then submit in the same run after approval.
  8. Preserve the browser when human approval or authentication completion is the only remaining blocker; otherwise delete it when the task is complete.
  9. Call complete_task exactly once, using success only for an achieved and verified result or failure for a blocked, incomplete, or failed task, then return the same terminal message.

Browser Execution Setup

  1. Make the Kernel browser execution tools available to the Openclaw Skills runtime, including manage_browsers, execute_playwright_code, computer_action, list_vault, fill_from_vault, and complete_task.
  2. Review the current Kernel documentation before implementation:
  3. Configure vault entries only when saved secrets or saved identity data are required. Supported setup kinds are login, payment, address, and phone.
  4. Start each task by creating a browser and keep the fast path bounded to approximately 90 seconds and six browser tool calls for uncomplicated work.

No package installation command is specified by the skill; use the current Kernel integration and documentation as the source of truth for runtime setup.

Browser Execution Data Schema & Taxonomy

Browser task state

Area Organization and rules
Browser session One browser is created and reused for the complete job; it is deleted after completion unless a human action or approval remains necessary.
Interaction methods execute_playwright_code handles navigation, inspection, extraction, and deterministic actions; computer_action handles visual or coordinate-level interaction.
Task result complete_task is called exactly once with success or failure, followed by the same terminal message.
Timing metadata The fast path targets 90 seconds and six browser calls; Playwright calls have a 30-second ceiling, locator waits are at most five seconds, and computer-action sleeps are at most two seconds.
Recovery state A blocked page permits up to two materially different approaches before reporting the verified blocker.

Vault taxonomy

  • Supported vault item kinds: login, payment, address, and phone.
  • request_vault_setup accepts kind, optional label, optional account, and target.
  • fill_from_vault supports username, password, cardholder_name, card_number, expiration, expiration_month, expiration_year, cvc, billing_postal_code, address, phone, identity, and token.
  • Raw secrets must never be placed in a normal browser call. After vault injection, the skill must not inspect filled values or capture screenshots that could expose them.
  • User-provided non-credential values can be used directly and do not need to be stored in the vault first.

Browser Execution Advanced Features

  • Secure opaque-handle vault integration for login credentials, payment cards, addresses, phone numbers, identities, and tokens.
  • Transaction approval gating that binds approval to the exact merchant, item, quantity, option, fees, and quoted total.
  • Automatic continuation after approval, including vault filling and submission in the same run.
  • Bounded blocked-page recovery with materially different tactics instead of repeated selector or code retries.
  • Dual interaction strategy combining Playwright automation with computer controls and final screenshots when visual confirmation is useful.
  • Human-in-the-loop preservation when authentication challenges, approval, or another manual action is required.
  • Strict timeout, wait, browser-call, and terminal-result controls for predictable Openclaw Skills execution.
  • Explicit distinction between verified success, incomplete work, and exact environmental blockers.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*