1
Prepare one clean audio segment
Pick a short phrase with a definite beginning and ending. Trim unwanted silence, but keep natural breaths and pauses that explain the performance. Listen for prominent consonants, especially closed-lip sounds such as p, b, and m. Those moments provide useful visual checkpoints. If you are filming, rehearse while playing the exact segment you intend to use; a different recording may have different pauses even when the words match.
2
Capture or prepare the face
For a filmed take, keep the speaker well lit, the mouth visible, and the camera steady. Play the audio while recording and perform the words rather than merely opening and closing your mouth. Record another take if the opening is late or a pause feels forced. For animation, prepare mouth poses and a face angle that makes them readable. Keep expressions subtle enough that the speech shapes remain visible.
3
Align, review, and export
Place the intended audio beneath the footage or animation. Match the first clear mouth movement to its sound, then check a closed-lip consonant near the middle and another cue near the end. If the beginning matches but the ending drifts, inspect playback speed and edits before shifting the entire clip. Review once with sound and once at reduced speed, then watch the exported file: the final render, not the timeline preview, is what viewers will see.