Use this skill whenever the user wants to convert text to speech, generate audio from text, create voiceovers, or produce spoken audio files. Triggers include: any mention of 'text to speech', 'TTS', 'read aloud', 'voice synthesis', 'generate speech', 'voiceover', 'narration audio', 'speak this text', or requests to turn written content into an audio/MP3 file. Also use when the user wants to pick a voice, adjust speech emotion/style, change speech rate or pitch, or compare TTS providers (Azure, Volcengine, Edge). If the user asks to 'make an audio version' of any text, 'record' a script, or produce '.mp3' output from text, use this skill. Do NOT use for speech-to-text, audio transcription, music generation, or sound effects.

Install

openclaw skills install @fengwm64/tts-api