LipSyncer

Talking Photo Maker — Make Any Portrait Speak with AI

Who this is for: For anyone who wants a still portrait to deliver a line: a birthday message from grandma's photo, a pet 'speaking' its thoughts, an avatar reading your script, or a creator photo shouting a hook into a Reel.

Upload a photo and 10 seconds of audio, and the face comes alive — mouth movements matched to the voice. 1 free video a day, no account, nothing installed.

Quick answer: A talking photo maker animates a still portrait to speak your audio: upload a photo, add up to 10 seconds of voice, and AI matches the mouth to the sound. The free tier gives one video a day — no account, no watermark, download the MP4 when it's done.

Choose a photoJPEG/PNG/WebP · a clear, front-facing portrait works best
Choose audioMP3/WAV/M4A/OGG · up to 10s · trim to the best moment

1 free video every day · no watermark · files are prepared locally in your browser

Core facts
InputOne photo (JPEG/PNG/WebP up to 10MB) + up to 10s of audio (MP3/WAV/M4A/OGG)
OutputMP4 video, 480p, water-mark free, mouth synced to your audio
ProcessingPhoto compression + audio conversion run locally in your browser
Free tier1 video/day, no account · upgrades via credit packs
Credits4 videos for $4.99 · 10 videos for $9.99 · 25 videos for $19.99 one-time, PayPal
StorageNothing stored — results are downloadable for 24 hours

How the talking photo maker works

A talking photo is a short video where a still image moves its mouth in sync with a voice track. Under the hood, an AI lip-sync model reads the face in your photo, follows the audio's timing — every phoneme, every pause — and renders new frames where the mouth (and a touch of the head and eyes) moves to match. The output is a normal MP4: it plays anywhere, posts anywhere.

On this page the whole pipeline runs like this: pick a photo, pick an audio clip (or record one), and both are prepared inside your browser before anything is sent. The photo is scaled down to 1024px and the audio is converted to a clean 16kHz WAV locally — then a single request goes to our server, which forwards it to the lip-sync engine and hands the finished MP4 back to your browser. Nothing is stored on our side, and the result appears right on the page with a download button.

The free tier gives you 1 talking video per day — full quality, no watermark, no sign-up. After that, credit packs keep it going: Starter — 4 lip sync videos ($4.99), Pro — 10 lip sync videos ($9.99), or Studio — 25 lip sync videos ($19.99). One credit is one video, no subscription.

Common uses

  • A photo of a pet 'saying' a birthday message for a family group chat
  • An old family portrait delivering a line you always wish it had said
  • A creator's photo announcing a channel milestone in a Short
  • A character or avatar still reading a script you recorded
  • A product mascot photo speaking a tagline

Talking photo vs. deepfakes — where the line is

This is a lip-sync tool, not a face-swap: the person in the output is the person in the photo, mouth moving to a voice you supply. That's why it's the right tool for pets, portraits, mascots and your own face — and the wrong tool for putting words into someone else's mouth. Use only photos you have the rights to, get consent for people who appear, and keep it clearly lighthearted. Our terms and copyright policy spell this out in plain language.

Frequently Asked Questions

What kind of photo works best?
A clear, front-facing portrait with the mouth visible and reasonably lit. Straight-on or slight angles both work; extreme profiles or sunglasses hiding the mouth give the model less to animate. Group photos are hit-or-miss — the model picks the dominant face. Photos up to 10MB are accepted; JPEG, PNG, WebP, GIF and BMP all work.
How long can the audio be?
10 seconds per video on every plan — enough for a sentence, a hook, or a chorus line. Longer audio is trimmed to the first 10 seconds automatically (you'll see a note when that happens). The audio can be MP3, WAV, M4A or OGG; it's converted to a clean WAV inside your browser before upload.
Do I need to record the voice myself?
No — any audio file works: a voice memo, a TTS clip from any tool, a song chorus, even a cartoon clip you have the rights to use. The engine only needs to hear the timing and tone of speech; where it came from doesn't matter.
Is there a watermark?
No. The MP4 you download is the full-quality render — no watermark, no badge, no signup gate. The free tier and paid credits produce the exact same file.
Is my photo or audio stored?
No. The photo is compressed and the audio is converted locally in your browser; both are passed through our server to the generation engine only to produce your video. We keep no copy afterwards. The finished video link stays live for 24 hours, so download it when it's ready.
Can I use someone else's photo?
Only with their consent. This tool is for your own photos, people who've agreed to appear, pets, characters you created, or assets you have the rights to. Using someone's face without permission — especially for anything they wouldn't endorse — is against our terms and, in many places, against the law.

More guides

How to Make a Lip Sync Video

From photo or clip to finished video: the whole process, with the settings and material choices that actually matter.

More related tools

Video Lip Sync

Upload a video, drop in new audio, and the mouth in the footage is re-synced to the new soundtrack — dub a clip, fix a take, or swap the lines.

Singing Photo

Make a photo sing: upload a portrait plus a song clip (up to 10 seconds) and the face performs it, mouth matched to the music.