Clawatar for Openclaw

Clawatar provides your AI agent with a fully interactive 3D VRM avatar body featuring lip-sync, expressions, and over 160 animations.

dongping-chen
v0.2.0
Feb 12, 2026
2
2.1k
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install clawatar

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install clawatar using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is Clawatar?

Clawatar is a powerful visual interface designed to bring AI agents to life through high-fidelity 3D models. By utilizing the industry-standard VRM format, this entry in the Openclaw Skills ecosystem allows developers to bridge the gap between text-based interactions and immersive visual experiences. It includes a web-based viewer that handles real-time rendering, skeletal animations, and automatic lip-syncing, transforming a standard LLM into a visible, speaking companion.

The system operates via a local WebSocket server, making it highly extensible and easy to control from any external script or agent. Whether you are building a virtual assistant, a VTuber-style character, or an interactive educational guide, Clawatar provides the technical foundation for realistic 3D representation and voice-driven interaction.

Clawatar Use Cases

  • Creating a visual persona for AI-driven customer service bots or personal assistants.
  • Building interactive VTuber-style characters controllable via an AI agent.
  • Implementing a standalone 3D VRM viewer with integrated voice chat capabilities.
  • Developing immersive AI companions for gaming or educational environments.
  • Prototyping multi-modal interfaces that combine text, voice, and 3D movement using Openclaw Skills.

How Clawatar Works

  1. The user launches the Clawatar web viewer which establishes a local WebSocket server at ws://localhost:8765.
  2. An AI agent or script sends structured JSON commands (such as play_action or speak) to the viewer via the WebSocket protocol.
  3. The viewer processes the command, triggering specific animations from a pre-loaded library or updating the character facial expressions.
  4. When a speak command is issued, the system utilizes ElevenLabs for high-quality TTS and synchronizes the avatar lip movements with the audio output.
  5. The interface provides real-time feedback, allowing users to interact directly via touch reactions or emotion bar controls on the web page.

Clawatar Setup

To get started with this project among your other Openclaw Skills, run the following commands:

# Clone and install
git clone https://github.com/Dongping-Chen/Clawatar.git ~/.openclaw/workspace/clawatar
cd ~/.openclaw/workspace/clawatar && npm install

# Start the Vite development server and WebSocket bridge
npm run start

Once running, navigate to http://localhost:3000. You will need to provide your own .vrm model by dragging and dropping it onto the page or configuring the model.url in clawatar.config.json. For voice features, ensure your ElevenLabs API key is configured in your environment variables.

Clawatar Data Schema & Taxonomy

Clawatar organizes its assets and configuration to ensure compatibility with various Openclaw Skills workflows:

Component Details
VRM Model User-provided 3D model file (.vrm).
Animations A library of 162 unique motion files stored in public/animations/catalog.json.
Configuration clawatar.config.json manages ports, default model paths, and voice settings.
WebSocket API Accepts JSON payloads with keys like type, action_id, name, weight, and text.
Expressions Built-in support for happy, angry, sad, surprised, and relaxed.

Clawatar Advanced Features

  • Real-time lip-syncing driven by advanced TTS analysis for natural speech visualization.
  • Extensive animation library including 162 distinct motions for greetings, thinking, dancing, and more.
  • Multi-scene support with environment presets like Sakura Garden, Night Sky, and Sunset.
  • Dynamic camera presets for Face, Portrait, and Full Body shots to enhance cinematic presentation.
  • Low-latency WebSocket control, allowing for seamless integration into complex Openclaw Skills automation pipelines.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*