A face that isn't yours, saying your words
Use it when: The content is direct address — advice, explanation, a pitch — and it reads wrong without a person delivering it.
- 01Write and produce the audio first. Lipsync is driven by the voice track, so the performance is decided before any face exists.
- 02Pick or generate the presenter's still and keep it. This image becomes the identity of the channel, and changing it later resets whatever recognition you'd built.
- 03Drive the face from the audio, then check the first and last two seconds — that is where sync drift shows up.
- 04Caption anyway. Synthetic delivery is harder to follow at low volume than a real one, and captions absorb most of that gap.
Run it here