ClawText Ingest provides production-ready memory ingestion for AI agents, supporting Discord, files, and URLs with automatic deduplication.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install clawtext-ingest
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install clawtext-ingest using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
ClawText Ingest is a powerful developer tool designed to bridge the gap between unstructured data sources and AI agent memory. By leveraging Openclaw Skills, developers can transform Discord forums, local files, web URLs, and JSON data into structured, deduplicated memories. This skill ensures that AI agents have access to high-quality, context-aware information through automatic YAML frontmatter generation and project-based routing.
Built for reliability and scale, this tool solves common challenges such as manual data entry, duplicate memories, and the loss of hierarchy in complex platforms like Discord. As a core component of the Openclaw Skills ecosystem, it integrates seamlessly with RAG layers, allowing agents to fetch and process knowledge autonomously through documented interaction patterns.
To get started with this addition to your Openclaw Skills library, install the package via npm or the OpenClaw CLI:
npm install clawtext-ingest
# OR
openclaw install clawtext-ingest
For Discord ingestion, set your DISCORD_TOKEN environment variable and use the following command to fetch forum data:
# Inspect forum structure
clawtext-ingest-discord describe-forum --forum-id YOUR_FORUM_ID
# Ingest with progress
clawtext-ingest-discord fetch-discord --forum-id YOUR_FORUM_ID
Finally, sync your changes to the memory cluster to enable RAG functionality:
clawtext-ingest rebuild
The skill organizes data into structured memory files with consistent metadata taxonomy. Below is the schema for generated memories:
| Field | Description |
|---|---|
| project | The unique identifier for the knowledge category (e.g., 'docs', 'research'). |
| type | The content classification (e.g., fact, decision, thread, adr). |
| entities | Auto-extracted keywords or linked concepts for agent retrieval. |
| date | ISO-8601 timestamp of when the content was ingested. |
| hash | SHA1 cryptographic hash used for 100% idempotent deduplication. |
Cross-session tracking is maintained in a .ingest_hashes.json file to ensure that Openclaw Skills can track processed items across different runs.
Loading
ClawSaver is an intelligent request-batching utility that reduces AI model API costs by 20–40% through efficient message buffering.

A comprehensive web automation skill for OpenClaw that enables professional headless browser interaction, data extraction, and visual page capture.

A specialized command-line interface for sending WhatsApp messages to third parties and managing message history.

A policy-driven firewall for AI agents that enforces spending limits and merchant restrictions before payments occur.

A professional workflow for researching marketplace trends and publishing high-quality skills to the ClawHub ecosystem.

A comprehensive verification toolkit for age-appropriate content filtering and regulatory compliance within the AI ecosystem.








































