← ClaudeAtlas

ai-voiceoverlisted

The AI narration / voiceover mini-skill (ElevenLabs-led). Use when someone wants an "AI voiceover," "narration," "text-to-speech for a video," "voice for my Reel/Short/explainer," "clone my voice," or to "dub a video into other languages." Picks the voice and model, writes for the ear, and directs the delivery; ElevenLabs generates the audio, the human mixes/reviews, WoopSocial schedules/publishes. Sits below the ai-video router, sibling to veo-3 and heygen. Consented voices only; disclose AI voice in ads/political.
social-media-skills/skills · ★ 5 · AI & Automation · score 73
Install: claude install-skill social-media-skills/skills
# ai-voiceover The **audio** producer of the video cluster — the counterpart to **veo-3** (scenes) and **heygen** (avatars) under the **ai-video** router. It picks the voice and model, writes for the ear, and directs the read; ElevenLabs renders the audio; a human mixes it in; WoopSocial schedules/publishes. ## The POV: 80% script + direction, 20% tool Most AI VO sounds robotic because people feed it **eye-written copy** and accept the **default read**. A great voiceover is mostly the script-for-the-ear and the direction. Write the way people talk, direct the delivery (model, Audio Tags, settings), and remember **social plays on mute** — so the VO supports captions, it doesn't carry the video alone. ## Read these first 1. **brand-profile** — audience, platform, non-negotiables. 2. **voice-builder** — the brand's **written** voice. This skill picks an **audio** voice + delivery that embodies it (keep them consistent). ## The framework: VOICE (Depth: `references/the-voice-framework.md`.) - **V — Voice match:** library / Voice Design / consented clone; fit brand + platform. - **O — Own the script for the ear:** spoken cadence, contractions, short sentences; read it aloud. - **I — Inflect & direct:** model by job (v3 expressive + Audio Tags / Multilingual v2 final / Flash draft); Stability ~0.3–0.5 expressive vs ~0.7–1.0 consistent; Similarity ~0.75–0.85; pronunciation. - **C — Caption alongside:** sound-off reality — VO supports captions; localize via Dubbing (70+ langs