ai-voiceoverlisted
Install: claude install-skill social-media-skills/skills
# ai-voiceover
The **audio** producer of the video cluster — the counterpart to **veo-3** (scenes) and **heygen**
(avatars) under the **ai-video** router. It picks the voice and model, writes for the ear, and
directs the read; ElevenLabs renders the audio; a human mixes it in; WoopSocial schedules/publishes.
## The POV: 80% script + direction, 20% tool
Most AI VO sounds robotic because people feed it **eye-written copy** and accept the **default
read**. A great voiceover is mostly the script-for-the-ear and the direction. Write the way people
talk, direct the delivery (model, Audio Tags, settings), and remember **social plays on mute** — so
the VO supports captions, it doesn't carry the video alone.
## Read these first
1. **brand-profile** — audience, platform, non-negotiables.
2. **voice-builder** — the brand's **written** voice. This skill picks an **audio** voice + delivery
that embodies it (keep them consistent).
## The framework: VOICE
(Depth: `references/the-voice-framework.md`.)
- **V — Voice match:** library / Voice Design / consented clone; fit brand + platform.
- **O — Own the script for the ear:** spoken cadence, contractions, short sentences; read it aloud.
- **I — Inflect & direct:** model by job (v3 expressive + Audio Tags / Multilingual v2 final / Flash
draft); Stability ~0.3–0.5 expressive vs ~0.7–1.0 consistent; Similarity ~0.75–0.85; pronunciation.
- **C — Caption alongside:** sound-off reality — VO supports captions; localize via Dubbing (70+ langs