Make a photo talk

Upload a portrait, add a voice you record, upload or type, and get a lip synced talking video. 2 credits a second, no watermark on any plan, and you own the result.

Make a talking photo

Real talking photo examples

Output from InfiniteTalk Fast, the lip sync model behind Percify's talking photos: one still image and one audio track in, each clip under 10 seconds. Press play for sound.

Check your photo first, free

The checker runs in your browser and tells you whether the face is big enough, facing the camera, lit and sharp. Nothing is uploaded until you choose to make the video.

Will this photo lip sync well?

Choose a photo and the checker finds the face, then looks at its size, angle, light and focus. It runs in your browser; the photo is not uploaded.

How it works

  1. Step 1

    Choose a photo

    A sharp, front facing head and shoulders photo of one person. A render or an illustrated face with clear eyes and mouth works too.

  2. Step 2

    Add the voice

    Record it, upload an audio file, or pick one of your saved voices or a public voice and type what it should say.

  3. Step 3

    Generate and download

    The mouth, jaw and head move with the audio. Download the MP4 with no watermark.

What it costs

Talking photos are priced per second of video. Runs that fail are refunded.

ModelCredits a second15 s clip60 s clip
InfiniteTalk Fast230120
InfiniteTalk, 480p460240
InfiniteTalk, 720p690360

Credit prices by plan are on the pricing page. Checked 17 September 2026.

Watermarks, rights and consent

  • No watermark on any plan, including the free one.
  • You own the content you generate, under Percify's terms. It is not exclusive: similar inputs can give similar output for someone else.
  • Use your own photo and voice, or ones you have permission to use.

Details, and how to check any other tool for hidden watermarks: AI avatar watermarks and commercial use.

Make it sound like you

For a video in your own voice, record a clean sample once and reuse it for every script. Check the recording with the free voice sample checker, then set up Clone Yourself with your photo and voice.

Make your photo talk

One photo and one voice in, a lip synced MP4 out. 2 credits a second, no watermark.

Make a talking photo

Frequently asked questions

Everything you need to know before you start.

Upload a clear, front facing photo, add the voice (record it, upload an audio file, or pick a voice and type what it should say), and generate. Percify animates the mouth, jaw and head to the audio and gives you an MP4 to download.

It depends on what you need: realistic lip sync on your own photo, your own cloned voice, no watermark, commercial use, or an API. Percify covers those on every plan and charges per second of video, but it is our product, so compare it on the same points with the other tools you are considering.

InfiniteTalk Fast costs 2 credits a second: a 15 second clip is 30 credits and a 60 second clip is 120. The higher quality InfiniteTalk model costs 4 credits a second at 480p or 6 at 720p. Runs that fail are refunded.

No. Percify adds no watermark on any plan, including the free one.

Yes. Percify’s terms say you own the AI generated content you create, subject to the terms. Output is not exclusive, and you need the rights to the photo and the voice you use.

Yes. Record or upload your voice for a clip, or save it and type new scripts for it. The free voice sample checker tells you if a recording is clean enough first.

A sharp, well lit, front facing head and shoulders photo of one person. The free photo checker on this page tests size, angle, tilt, light and focus before you spend anything.

Your own, or a photo of someone who agreed to it. Percify’s terms require that you have the rights to what you upload.