dialogue-and-voicelisted
Install: claude install-skill raphaelbgr/ai-video-prompting
# Dialogue, voice and audio
Audio is generated whether you direct it or not. Leaving it blank is the single most
common mistake: the model invents speech, usually in English, and burns the clip.
---
## 1. The audio hierarchy
Direct these four, in this order of priority:
1. **Dialogue** — foreground, most important
2. **SFX** — tied to a visible action
3. **Ambience** — the background bed
4. **Music** — **add in post-production only** (see §6)
---
## 2. Dialogue format
```
Dialogue in <LANGUAGE>: <Speaker> speaks in <tone/delivery>. He/She says: <exact words>
Voice: <gender>, <age> years old, <language + accent>, <delivery>, <energy>, <emotional quality>.
```
**Three rules that each cost a generation to learn:**
### Quotation marks — the contested one
> `[CONTESTED]` Google's Veo prompt guide shows attributed dialogue **with** quotation
> marks (*A woman says, "We have to leave now."*). Separate Google guidance says to use
> **a colon after the speaker's action and avoid quotation marks**, because quotes push
> the model toward rendering the text visually in the video.
>
> **Field testing sided with the no-quotes form**, and it composes with the known caption
> bug — anything that nudges the model toward rendering text makes burned-in subtitles
> more likely. **Default to the colon form.** If speech fails to trigger at all, try the
> quoted form as a fallback and note which one your footage came from.
```
Prefer: Man says: Bom dia!
Fallback: A woman says, "We hav