Hey there! I had a question about lip syncing. As Sound Director, Iâm currently filling out an x-sheet with phonemes to send to my director, but my God is it tedious. My process consists of having the animatic open, scrolling through a couple frames, then filling the sheet out in Excel. Do you have any tips? Am I just missing out an obvious way to make it easier? Or do I just power through it? Thank you!!
Oh my gosh, what a great question to get in my inbox! Thanks for writing in!
As for lip assignment.... ugh. I feel your pain! To be honest the real secret is âkeep doing it until you get betterâ but tips I can think of off the top of my head to do it better/faster are:
Work with a mirror in a comfortable spot so you can say the words out loud and see your own mouth shape easily. Have your mouth chart up next to the mirror so you can compare your mouth to the chart mouths  (spread them out if itâs more than one page!) at a glance.
If youâre not sure what shape your mouth is making try to over-emphasize the word like crazy and âfeelâ what shape your mouth is making.
If possible scrub through your audio at half speed sometimes to hear the individual sounds
Break up the work into smaller chunks! For example:
Letâs say the line is âHello my name is Bob.â
Your timing sheet should have the frames of animation as your unit of measure. So do some rough timing of when each word/sound starts and stops in your track! You CAN use a stopwatch for this but personally I think what I describe below is much easier.
I donât know what youâre working with, but letâs say you have an MP3 of your audio track (or you can extract an mp3 from your animatic). Do that, and then open it in a program like Audacity (which is free) or a video editor that lets you count things in frames and shows waveform data.
If you look at the audio as a waveform you can zoom in and easily find the spots where the sounds happen! So for my imaginary âhello my name is Bob.â assignment, unless the file has tons of music or other voices in it, you can look at the waveform, and the first spike in the track will be when the word âhelloâ starts being said. When the spike of activity ends, theyâre done saying hello. Second spike is âmyâ, etc...
So make some marks on those rows of your sheet, to indicate where each word starts and stops (or make notes on paper for yourself but tbh itâs best to mark up the sheet itself!). Fill in all the spots between sounds with the letter or number your team is using to represent the closed mouth shape. You may have to tweak this later to get this just right, but as a starting point it will help you identify when there IS and ISNâT sound.
Then, on the first row of each word, on the margins or on the side (wherever youâve got free space) write the word phonetically - pay attention to any weird accent your character may have! Do they say ah-lum-in-um or a-loo-min-e-um? Â Do they say âHellooooooooooooooâ in a drawn out way or âHello!â in a fast way?
Then Iâm afraid itâs just listening carefully to the track - scrubbing the spiked zones to identify what frame number it is when the H sound ends and E sound begins, E ends and L begins, L ends and O starts, etc etc.
As you practice, you get better and faster at this and you need the chart less, need the mirror less! But until then good luck and godspeed :) Please write back and let me know if this helped!










