Voice Text to Meme Skill for Openclaw

A specialized AI skill that transforms voice-to-text input into expressive, chat-ready meme images based on sentiment and tone.

hei-maom
v1.0.0
Mar 14, 2026
0
774
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install voice-text-to-meme

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install voice-text-to-meme using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is Voice Text to Meme Skill?

The Voice Text to Meme skill is a creative tool designed to bridge the gap between spoken thoughts and visual communication. It takes raw or polished voice-recognized text and intelligently generates a single, high-quality meme or sticker suitable for messaging platforms. By analyzing the intent and tone of the input, the skill selects appropriate visual styles—ranging from cute stickers to exaggerated internet memes—to ensure the visual matches the vibe of the conversation.

As a standout among Openclaw Skills, it prioritizes ease of use and professional visual output. It leverages the doubao-seedream model to render text directly onto images or provide clean templates for further customization. This ensures that users can communicate with humor and personality without needing graphic design skills.

Voice Text to Meme Skill Use Cases

  • Creating instant reaction memes from voice messages in fast-paced group chats.
  • Generating custom team stickers to express speechlessness, excitement, or sarcasm.
  • Refining long-winded voice transcriptions into punchy, viral-style meme captions.
  • Producing clean visual templates for developers who want to overlay custom UI captions.

How Voice Text to Meme Skill Works

  1. The skill receives input text, prioritizing polished versions over raw voice-to-text for better context.
  2. It analyzes the text's mood to classify it into categories like speechless, aggrieved, celebratory, or sarcastic.
  3. The AI generates a concise caption (typically under 12 characters) and a detailed visual prompt for the image generator.
  4. Based on user settings, it chooses between 'direct-text' (text baked into the image) or 'template' (no text) modes.
  5. It executes a Python script to call the image generation model using OpenAI-compatible APIs.
  6. The resulting image file path and optional metadata are returned for immediate sharing.

Voice Text to Meme Skill Setup

To get started with this skill, configure your environment variables and ensure the generation script is accessible:

# Set your API credentials
export MEME_MODEL_API_KEY="your_secret_key"
export MEME_MODEL_BASE_URL="https://models.audiozen.cn/v1"
export MEME_MODEL_NAME="doubao-seedream-4-5-251128"

# Run the generation script manually if needed
python scripts/generate_meme.py --text "I can't even" --mode direct-text --size 2K

Voice Text to Meme Skill Data Schema & Taxonomy

The skill processes text inputs and environment configurations to output image assets. Below is the data structure overview:

Attribute Type Description
original_text String The raw output from voice recognition.
polished_text String Refined text used as the primary source for the meme.
style Enum Visual direction (e.g., Cute Sticker, Exaggerated Meme, Minimalist).
output_dir Path The local directory where generated memes are stored.
mode String Specifies if text is rendered on the image or provided as a template.

Voice Text to Meme Skill Advanced Features

  • Automatic Tone Detection: Intelligently switches visual styles based on 8+ distinct emotional profiles.
  • High-Resolution Rendering: Supports up to 2K resolution for crisp, professional-looking assets.
  • Intelligent Caption Compression: Automatically shortens long sentences into punchy, high-impact meme copy.
  • Flexible Integration: Compatible with any OpenAI-style image API for easy model swapping.
  • Safety Filtering: Avoids generating offensive or low-quality content in professional chat environments.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*