aliyun-tts for Openclaw

An Alibaba Cloud-powered Text-to-Speech synthesis skill for generating high-quality audio from text strings.

guang384
v1.0.0
Jan 29, 2026
2
4.9k
0

Install & Download

1. ClawHub CLI

The fastest way to install a skill directly from the registry.

npx clawhub@latest install aliyun-tts

2. Manual Installation

Copy the skill folder to one of these locations

Global
~/.openclaw/skills/
Workspace
<project>/skills/

Priority: Workspace > Local > Bundled

3. Prompt Installation

Copy this prompt to OpenClaw to install it automatically.

Help me install aliyun-tts using Clawhub. If Clawhub is not installed, install it first (npm i -g clawhub).

Prefer to download?

Get the raw skill files in a ZIP archive.

What is aliyun-tts?

The aliyun-tts skill provides a professional interface for the Alibaba Cloud Text-to-Speech (TTS) synthesis service. It allows developers and AI agents to convert plain text into natural-sounding speech across various voices, languages, and audio formats. By integrating this capability into your library of Openclaw Skills, you can create more immersive and accessible user experiences.

This skill is designed to handle everything from basic voice notifications to complex chat-based voice replies. It abstracts the Alibaba Cloud API complexity into a simple command-line interface, making it easy to generate MP3 or WAV files on the fly for any application or automated agent response.

aliyun-tts Use Cases

  • Creating automated voice replies for AI chat agents.
  • Generating localized audio content for global applications using regional voices.
  • Developing accessibility tools that read text aloud for visually impaired users.
  • Building notification systems that provide audible alerts in smart home or office environments.
  • Batch processing text files into audio for podcasts or educational materials.

How aliyun-tts Works

  1. The skill receives a text input and optional parameters such as voice name or sample rate.
  2. It authenticates with the Alibaba Cloud Intelligent Speech Interaction service using the configured App Key and Access Keys.
  3. The text is sent to the Alibaba Cloud synthesis engine which processes the natural language and converts it to audio data.
  4. The resulting audio stream is downloaded and saved to a local file (defaulting to tts.mp3).
  5. If used within a chat context, the file path is returned using the standard media protocol for playback.

aliyun-tts Setup

Prerequisites

You must have an active Alibaba Cloud account and the required API credentials (App Key, Access Key ID, and Access Key Secret).

Configuration

It is recommended to use the CLI configuration to securely store your credentials in your Openclaw Skills environment:

# Configure App Key
clawdbot skills config aliyun-tts ALIYUN_APP_KEY "your-app-key"

# Configure Access Key ID
clawdbot skills config aliyun-tts ALIYUN_ACCESS_KEY_ID "your-access-key-id"

# Configure Access Key Secret
clawdbot skills config aliyun-tts ALIYUN_ACCESS_KEY_SECRET "your-access-key-secret"

aliyun-tts Data Schema & Taxonomy

Property Type Description
Output Path File Path The location where the generated .mp3 or .wav file is stored.
Voice String The identifier for the synthesized voice (e.g., siyue, xiaoxuan).
Format String The audio encoding format (default: mp3).
Sample Rate Integer The frequency of the audio in Hz (default: 16000).
Media Tag Protocol The MEDIA:/path/to/file syntax used for agent integration.

aliyun-tts Advanced Features

  • Support for multiple high-quality voices including siyue, xiaoxuan, and xiaoyun for diverse tonal needs.
  • Customizable audio parameters including sample rate and file format to match specific playback hardware.
  • Direct integration with chat agents via the MEDIA protocol for real-time voice response generation.
  • Simple CLI-based execution that can be easily incorporated into larger automation scripts.

SKILL.md


Loading

Related Openclaw Skills

METADATA

Github Stars: 0
forks: 0

Featured*