A bridge that enables AI agents to capture and process real-time audio streams from any browser tab without platform-specific integrations.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install browser-audio-capture
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install browser-audio-capture using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
Browser Audio Capture is a specialized utility designed to give AI agents auditory capabilities within a web environment. By acting as a bridge between the browser's media layer and your AI pipeline, it allows for the seamless extraction of audio from meetings, webinars, and streaming platforms. As part of the Openclaw Skills ecosystem, this tool eliminates the need for complex API keys or per-platform OAuth setups, making it a universal solution for developers building intelligent listeners.
This skill is built for flexibility, supporting both a Python-based CLI for automated workflows and a dedicated Chrome extension for manual, persistent capture. Whether you are building a real-time meeting assistant or an automated podcast note-taker, this skill provides the raw audio data necessary to power modern transcription and analysis models.
To get started, launch Google Chrome with the remote debugging port enabled:
/Applications/Google\ Chrome.app/Contents/MacOS/Google\ Chrome \
--remote-debugging-port=9222 --user-data-dir=$HOME/.chrome-debug-profile &
Install the required Python dependencies:
pip install aiohttp
For the CLI experience, use the following commands to manage your capture sessions:
# Detect and capture audio from a meeting tab
python3 -m browser_capture.cli capture
# Watch for new tabs every 15 seconds
python3 -m browser_capture.cli watch --interval 15
To use the Chrome Extension, navigate to chrome://extensions/, enable Developer mode, and use 'Load unpacked' to select the extension directory.
The skill streams data as a JSON payload to the configured endpoint. This structured format ensures that the Openclaw Skills consumer receives all necessary context for the audio stream.
| Property | Type | Description |
|---|---|---|
sessionId |
String | A unique identifier for the specific capture session. |
audio |
String | Base64 encoded PCM16 audio data. |
sampleRate |
Integer | The sample rate of the stream (default 16000). |
format |
String | The audio format (e.g., "pcm16"). |
tabUrl |
String | The source URL where the audio is being captured. |
tabTitle |
String | The HTML title of the source tab. |
Loading
An automated health monitoring and recovery watchdog designed to protect OpenClaw gateways from faulty configuration changes.

ShieldCortex provides AI agents with a secure, persistent memory system featuring semantic search, knowledge graphs, and a 6-layer defense pipeline.

Search and find award flight availability across 24 mileage programs using the seats.aero partner API.

The Mermaid Diagrams Skill enables the programmatic generation of visual charts and diagrams directly from text-based Mermaid syntax.

An always-on ambient intelligence layer that builds a continuous knowledge graph of your conversations and projects without needing explicit commands.

Ambient audio capture and local transcription for AI agents via wearable devices like Omi and Apple Watch.








































