Install
openclaw skills install @social-media-skills/ai-voiceoverThe AI narration / voiceover mini-skill (ElevenLabs-led). Use when someone wants an "AI voiceover," "narration," "text-to-speech for a video," "voice for my Reel/Short/explainer," "clone my voice," or to "dub a video into other languages." Picks the voice and model, writes for the ear, and directs the delivery; ElevenLabs generates the audio, the human mixes/reviews, WoopSocial schedules/publishes. Sits below the ai-video router, sibling to veo-3 and heygen. Consented voices only; disclose AI voice in ads/political.
openclaw skills install @social-media-skills/ai-voiceoverThe audio producer of the video cluster — the counterpart to veo-3 (scenes) and heygen (avatars) under the ai-video router. It picks the voice and model, writes for the ear, and directs the read; ElevenLabs renders the audio; a human mixes it in; WoopSocial schedules/publishes.
Most AI VO sounds robotic because people feed it eye-written copy and accept the default read. A great voiceover is mostly the script-for-the-ear and the direction. Write the way people talk, direct the delivery (model, Audio Tags, settings), and remember social plays on mute — so the VO supports captions, it doesn't carry the video alone.
(Depth: references/the-voice-framework.md.)
Eleven v3 (expressive, Audio Tags) or Multilingual v2 (polished long-form) for finals;
Flash/Turbo for drafts/real-time at ~half the credits. Draft on Flash, render finals on
v3/Multilingual v2. Full capabilities/pricing: references/elevenlabs-2026-capabilities.md; worked
scripts: references/script-for-the-ear-and-recipes.md.
references/consent-disclosure-and-tools.md.)Router: ai-video. Sibling producers: veo-3 (scenes), heygen (avatars).
captions-and-clipping pairs VO with sound-off captions + long→Short cuts. VO feeds
reels-script, youtube-shorts, youtube-long-form, linkedin-growth,
cross-platform-repurposing. Connection: tools/integrations/elevenlabs.md (+ tools/REGISTRY.md).
Publish: scheduling-and-queue → WoopSocial.
A voice + model chosen for the job and brand; a script written for the ear; delivery directed (tags/ settings/pronunciation); sound-off captions planned and localization handled where needed; consent verified and AI disclosure planned; the generate→mix/review→publish chain routed to scheduling-and-queue → WoopSocial; no unconsented cloning, no fabricated metrics.