AI Lip Sync
This ai lip sync tool runs Sync Lipsync 2 through fal on a video with one clearly visible speaking face plus your own replacement speech (MP3, WAV or M4A, up to 60 seconds and 20 MB). Video and speech must match in length within 0.25 seconds — there's no automatic looping — and you must confirm you have permission for the face, video and voice. Pricing is 12.5 credits per second of video; a 5.04-second clip runs 64 credits (rounded up). Sign in and upload both files; the server measures them and shows that exact quote, valid for 5 minutes, and only charges once you confirm. Results are private MP4s listed under Recent jobs (latest 20); a provider or output-validation failure refunds automatically, and you can leave and recover the job after signing in again. A Sept 27, 2026 test matched a 5-second portrait to a synthesized voice — mouth shapes followed the speech and closed during the trailing silence, but always review lip timing and audio before you use a result. Never use this to impersonate a real person without consent.
Updated Sep 9, 2026
Sync your video to speech
Upload video and speech → confirm permission → review the quote → download MP4.
Checking availability…
Sign in to upload and recover your resultsSync Lipsync 2 · Use one clearly visible speaking face. Trim speech and video to matching lengths (within 0.25 seconds). No automatic looping. Output requires visual and listening checks.
3. Your result
Your processed video will appear here. Recent jobs are saved to your account below.
No verified lip-sync sample is published yet. Paid access stays disabled until output acceptance.
Workflow and limitations
Two files, matched to within a quarter second
Upload a video up to 60 seconds with one clearly visible speaking face, plus replacement speech as MP3, WAV or M4A — up to 60 seconds and 20 MB. Video and audio lengths must match within 0.25 seconds; there's no automatic looping, so trim before you upload.
12.5 credits per second, quoted exactly
Pricing is 12.5 credits per second of video — a 5.04-second clip runs 64 credits (rounded up). The server measures both files and shows the exact quote, valid for 5 minutes, and only charges once you confirm.
Evidence: mouth shapes tracked a synthesized voice
A Sept 27, 2026 test matched a 5-second portrait to a synthesized voice; mouth shapes followed the speech and closed during the trailing silence. Always review lip timing and audio yourself before you use a result.
How to use it
- 01
Sign in, then upload a video with one visible speaking face and your replacement speech (MP3, WAV or M4A). Confirm you have permission for the face, video and voice.
- 02
Review the exact quote — 12.5 credits per second of video, valid for 5 minutes — then confirm.
- 03
Download the new MP4 and check lip timing and audio, or reopen the job later from Recent jobs.
Prompt ideas
Untested prompt drafts. For recorded outputs, use the playable examples and testing record.
“A 5-second clip of one person facing the camera, mouth clearly visible and unobstructed.”
“Replacement speech trimmed to match your video's length within 0.25 seconds, with no background music.”
“A steady, well-lit close-up rather than a wide shot, so the mouth stays large enough to track.”
FAQ
Evidence, sources and testing scope
Local tests cover owned video/audio measurement, quotes, billing safety and stored media delivery using injected provider responses. No actual lip-motion or speech-quality result is verified; paid access remains disabled.
Cicadas.ai testing record · Reviewed 2026-09-21
See samples, methods, costs and limitationsRelated tools
AI Video Generator with Sound
Create a 4–8 second Veo 3.1 Fast MP4 with an audio track. Describe action, ambience or dialogue, review exact credits, then recover and download it.
Veo 3.1 Fast
Kling 3.0 Pro · Recorded Cicadas sampleAnimate a Photo with AI
Upload one portrait, pet or landscape photo, choose a restrained movement and generate a silent 5- or 10-second Kling MP4 with an exact credit quote.
Kling 3.0 Pro