KlingAIResearch
LivePortrait
Animates a face photo or video with motion from another clip
LatentSync re-syncs a talking-face video to a new audio track. ByteDance's open-source lip-sync model uses diffusion (the method behind Stable Diffusion). It ships with a Gradio app, a command-line script, training code and eval tools. Inference needs 8-18 GB of VRAM, and you download the checkpoints yourself.
Best for: Researchers and tinkerers who want to dub talking-head or anime clips with a diffusion lip-sync model.
Rewritten from the README. Check the repo for the latest steps.
source setup_env.sh to install packages and download the checkpoints.python gradio_app.py../inference.sh.inference_steps (20-50) for sharper video, or guidance_scale (1.0-3.0) for tighter lip-sync../data_processing_pipeline.sh, then run ./train_unet.sh or ./train_syncnet.sh.Other editing & audio projects people compare with LatentSync.
KlingAIResearch
Animates a face photo or video with motion from another clip
MeiGen-AI
Lip-synced video or photo from new audio, any length
KlingAIResearch
Tool that redraws a video character's mouth to match a new audio track
Huanshere
Translated subtitles and a dubbed voice track for any video
Do you maintain LatentSync? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/bytedance-latentsync)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. Last push 2025-06-20, outside our 90-day window. We kept it because it's still widely used. Built on AnimateDiff; borrows code from MuseTalk, StyleSync, SyncNet and Wav2Lip. We write these descriptions ourselves. The repo's own docs are the final word.