MenuVision transforms restaurant menus from any source into interactive, visually rich HTML digital menus using advanced AI vision and image generation.
The fastest way to install a skill directly from the registry.
npx clawhub@latest install menuvision
Copy the skill folder to one of these locations
~/.openclaw/skills/ <project>/skills/ Priority: Workspace > Local > Bundled
Copy this prompt to OpenClaw to install it automatically.
Help me install menuvision using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).
Get the raw skill files in a ZIP archive.
MenuVision is a comprehensive technical pipeline designed to bridge the gap between physical restaurant menus and modern digital experiences. By leveraging Gemini Vision for data extraction and AI image generation for food photography, it automates the creation of professional HTML menus. This skill is part of the Openclaw Skills collection, providing a robust solution for developers to convert unstructured URLs, PDFs, or photographs into structured JSON data and subsequently into a polished, responsive web interface.
The tool is specifically engineered for high-fidelity extraction, handling complex multi-lingual menus (including CJK scripts) and diverse price formats. It produces a feature-rich output that includes Instagram-style grids, interactive selection receipts, and client-side currency conversion, making it an ideal choice for digital signage, online ordering previews, or restaurant digitization projects.
To get started with this skill from the Openclaw Skills library, ensure you have Python 3.9+ installed and configure your environment:
# Install core dependencies
pip install google-genai Pillow requests beautifulsoup4 PyMuPDF
# Install browser automation for JS-heavy sites
pip install playwright && playwright install chromium
# Configure your Google API Key
export GOOGLE_API_KEY="your_api_key_here"
The skill operates on a strict JSON data contract to ensure pipeline reliability. Below is the primary structure:
| Component | Key Fields | Purpose |
|---|---|---|
| Restaurant | name, cuisine, tagline |
Defines header branding and image generation context. |
| Sections | title, category, items |
Groups menu entries; category drives image generation logic. |
| Items | code, name, price, dietary |
Individual menu entries with pricing and allergen metadata. |
| Allergen Legend | code: display_name |
Maps technical codes to human-readable allergen names. |
| Metadata | languages, currency |
Controls CJK font loading and currency display symbols. |
Loading
A comprehensive tool for generating high-fidelity images and cinematic videos using xAI's Grok Imagine Extended models.

A daily optimization engine that delivers actionable strategies to improve AI agent performance and reduce operational costs.

A robust automated backup and disaster recovery system designed to preserve OpenClaw agents, skills, configurations, and long-term memory.

An MPC-based wallet infrastructure for AI agents to securely sign blockchain transactions without exposing private keys.

Generate single-use virtual credit cards on the fly to complete secure online payments during automated agent tasks.

A suite of automation skills that allow AI agents to control macOS via Accessibility APIs, handling everything from window management to OCR-based interactions.








































