Generate the voiceover MP3 AND derive word-level subtitle timings for a short-form video job. Calls ElevenLabs' timestamped endpoint, saves audio/voiceover.mp3, computes word boundaries from character alignment, and rewrites input.json's subtitles[] to match the actual voiceover timing. Use whenever the orchestrator hands off voiceover generation.

Install

openclaw skills install @pushpendrachauhan/elevenlabs-tts