tts-whatsapp
Send high-quality text-to-speech voice messages on WhatsApp in 40+ languages with automatic delivery
Why use this skill?
Send high-quality, AI-generated voice messages on WhatsApp in 40+ languages. Automate broadcasts, support groups, and enhance communication with OpenClaw.
Install via CLI (Recommended)
clawhub install openclaw/skills/skills/hopyky/tts-whatsappWhat This Skill Does
The tts-whatsapp skill is a powerful automation tool that integrates text-to-speech (TTS) capabilities directly with WhatsApp messaging. It leverages the Piper TTS engine to generate high-quality, natural-sounding voice messages from text input. Once the audio is generated, the skill automatically transcodes the file into the OGG/Opus format required by WhatsApp and utilizes the Clawdbot agent to deliver the message to any individual contact or group. This tool eliminates the need for manual voice recording, allowing users to broadcast information, updates, or personal messages in over 40 languages with extreme speed and efficiency.
Installation
To get started, ensure you have the necessary system dependencies. First, install Piper TTS via pip3 install --user piper-tts. Next, install FFmpeg, which is required for audio format conversion—use brew install ffmpeg on macOS or apt install ffmpeg on Linux. Download your preferred voice models from the official Rhasspy Hugging Face repository and move them to ~/.clawdbot/skills/piper-tts/models/. Finally, install the skill itself by running the command clawhub install openclaw/skills/skills/hopyky/tts-whatsapp in your terminal.
Use Cases
This skill is perfect for scenarios where a human touch is needed without the effort of real-time recording. Businesses can use it for automated status notifications or appointment reminders in the customer's native language. Content creators can quickly narrate scripts for group updates. It is also highly effective for accessibility, enabling users to send messages in languages they may not speak fluently, or for users who prefer listening to messages rather than reading them in high-noise environments.
Example Prompts
- "tts-whatsapp 'Your scheduled maintenance is confirmed for tomorrow at 10 AM' --target '+447700900123'"
- "tts-whatsapp 'Bonjour, voici le compte rendu de la réunion' --lang 'fr_FR' --voice 'siwis' --target '[email protected]'"
- "tts-whatsapp 'The server update is complete, everything is back online' --target '+15550199' --quality 'high'"
Tips & Limitations
To maintain performance, keep messages concise; the system averages around 2-3 seconds for a full delivery cycle. If you are sending to a large group, ensure your group ID is correctly retrieved. Note that the quality of the audio is highly dependent on the voice model chosen—'high' quality settings will increase generation time slightly but provide superior clarity. Always ensure your environment variables in clawdbot.json are set correctly to avoid repetitive typing of common target numbers.
Metadata
Not sure this is the right skill?
Describe what you want to build — we'll match you to the best skill from 16,000+ options.
Find the right skillPaste this into your clawhub.json to enable this plugin.
{
"plugins": {
"official-hopyky-tts-whatsapp": {
"enabled": true,
"auto_update": true
}
}
}Tags
Flags: network-access, file-write, file-read, external-api
Related Skills
narrator-ai-cli
Create AI-narrated film/drama commentary videos via CLI. Two workflow paths (Original & Adapted narration), 100+ movies, 146 BGM tracks, 63 dubbing voices in 11 languages, 90+ narration templates. Use when creating narration videos, film commentary, short drama dubbing, or video production.
narrator-ai-cli
AI电影解说视频自动生成技能(AI解说大师 CLI Skill)。当用户需要创建电影解说视频、短剧解说、影视二创、AI配音旁白视频、film commentary、video narration、drama dubbing、movie narration时触发。内置93部电影素材、146首BGM、63种配音音色(11种语言)、90+解说模板。通过narrator-ai-cli命令行工具实现:搜片选片→选择模板→选BGM→选配音→生成文案→合成视频的全流程自动化。CLI client for Narrator AI (AI解说大师) video narration API. Use when user needs to create AI narration videos, manage narration tasks, browse dubbing/BGM/material resources, or automate video production.
podcast-agent
Search articles on any topic, generate a two-host dialogue script, and synthesize podcast audio via TTS. Turn long reads into listenable content.
ressemble
Text-to-Speech and Speech-to-Text integration using Resemble AI HTTP API.
ym-mediatoolkit
流式视频处理工具集 - 压缩、封面提取、音频转换,无需下载完整视频