How to make a talking avatar video

Create a talking AI avatar from a single photo. Animate a presenter, add a voice, and lip-sync the words for content, training, and presentation videos in Morphic.

To make a talking avatar video, upload a clear presenter photo to Morphic, animate the face with subtle motion, add a voice from your script, and lip-sync the words so the avatar appears to speak. Attaching the same photo as a reference keeps the presenter consistent across videos, so you build a repeatable on-screen spokesperson without booking a shoot.

Steps at a glance

  1. Choose a clear presenter photo
  2. Animate it as a talking presenter
  3. Add the script and voice
  4. Lip-sync the words to the face
  5. Reuse the avatar and export

Build your talking avatar step by step

1.

Choose a clear presenter photo

Start with a sharp, front-facing photo of the person who will present, well lit and looking toward the camera against a simple background. This one image becomes your presenter, so pick it deliberately and treat it as a casting decision, not a throwaway. Attach it as a reference from the start, because the reference is what holds the face steady across every clip and turns a set of separate videos into one recognisable spokesperson rather than a slightly different person each time.

2.

Animate it as a talking presenter

Animate the photo with the subtle, human motion a presenter has on camera: small head movement, natural blinking, a slight smile, an engaged expression. Keep it understated, because big head turns or exaggerated motion let the face drift from the original. The goal is a presenter who looks alive and attentive while the voice and lip-sync, added next, carry the actual delivery of the message. Overdoing the motion here is the most common mistake, and it is what makes a talking avatar tip from convincing into uncanny.

3.

Add the script and voice

Write what the presenter should say and generate a voice for it, directing the tone and pacing so it sounds like a person presenting rather than a script being read. Pick one voice and keep it for the whole series so the presenter sounds consistent. If you prefer, convert a recording into a different voice with the voice changer, noting that cloning your exact voice from a sample is not yet available.

4.

Lip-sync the words to the face

Combine the animated presenter and the voice with lip-sync so the mouth matches the speech. Clear front-facing footage and clean audio give the most accurate result, so keep the face unobstructed and avoid background noise in the voice track. If the sync drifts, re-run it with a cleaner audio take. The AI lip-sync guide covers the settings in more detail.

5.

Reuse the avatar and export

For a single video, review the resemblance and sync, then export. For a series, reuse the same reference photo and voice on every new script so the presenter stays identical, and assemble longer pieces on the Compose timeline with captions and any b-roll cut in over the presenter. This is what makes an avatar worth the setup: once the photo and voice are dialed in, every future video is just a new script through the same presenter, so a whole training course or content series shares one consistent face. Free exports carry a watermark; a paid plan exports clean and at higher resolution, which matters for training and published content people watch closely.

Talking avatar versus filming a presenter

A talking avatar replaces the shoot, not the message. Here is the trade-off, Morphic first but not the only route.

AI talking avatarFilm a presenterStock presenter footage
You provideOne presenter photo + a scriptTalent, camera, and a studioA licence and a search
Update the scriptRegenerate the videoReshoot with the talentNot possible; find new footage
ConsistencySame face via a referenceDepends on re-booking talentDifferent person each clip
Best forContent, training, explainersHigh-end brand filmsGeneric filler shots

If the presenter is you specifically, how to make an AI video of yourself covers animating your own photo, and adding AI voiceover goes deeper on directing the voice the avatar speaks with.

Put it together

A talking avatar is one photo, subtle motion, a directed voice, and accurate lip-sync, held consistent by a reference. Choose the presenter photo carefully, keep the motion understated so the face holds, and reuse the same reference and voice across every script. Do that and you have a repeatable on-screen presenter you can point at any message, no shoot, no studio, no talent to re-book.

FAQ

What is a talking avatar video?
It is a video where a presenter appears to speak to camera, built from a single photo rather than filmed. You animate the face with subtle motion, add a voice from a script, and lip-sync the words so the avatar delivers your message without a shoot.
Do I have to be on camera?
No. That is the point of a talking avatar: you provide one clear photo of the presenter, and Morphic animates it and syncs a voice to it. It suits creators who want a consistent on-screen presence, and teams making training or explainer video without booking talent and a studio.
Can I reuse the same avatar across videos?
Yes. Attach the same presenter photo as a reference on every video so the face and look stay consistent from clip to clip. That is how you build a recognisable presenter across a series rather than a slightly different person each time you generate.
Does the avatar need my own voice?
No. Generate a voice you direct for tone and pacing, or convert a recording into a different voice with the voice changer. Cloning your exact voice from a sample is not yet available, so pick a generated voice that fits the presenter and keep it consistent across the series.
Is it free to make a talking avatar?
Morphic has a free plan you can start on, so you can build a talking avatar without paying. Free exports carry a watermark, and a paid plan removes it and unlocks longer, higher-resolution work. Test the avatar and voice free, then upgrade to export clean for publishing.