MeiGen-AI
InfiniteTalk
Lip-synced video or photo from new audio, any length
- Wan 2.1-I2V-14B-480P
- chinese-wav2vec2-base
- InfiniteTalk
- +3
X-Dub takes a video plus an audio file and redraws the mouth to match. It's the official code for a research paper, and the public build runs on Wan2.2-TI2V-5B. Budget about 21 GB of VRAM, and it handles one person per clip. The public release is weaker than the paper version: flicker, drift (the face slowly changing) and noisy frames.
Best for: Re-dub artists who need to match one on-screen character's mouth to a new audio track.
ref_cfg_scale, audio_cfg_scale and 25-50 stepsRewritten from the README. Check the repo for the latest steps.
conda activate x-dubgit clone https://github.com/KlingAIResearch/X-Dub.gitconda create -n x-dub python=3.10 -ypip install -e . --no-depspip install -r requirements.txtcheckpoints/: the command belowhf download KlingTeam/X-Dub --local-dir ./checkpoints --repo-type modelmkdir -p dwpose_tools/models then the command belowcp -r ./checkpoints/dwpose_tools/models/. ./dwpose_tools/models/python infer_lip_sync_pipeline.py --video_path assets/examples/video.mp4 --audio_path assets/examples/audio.wav --ckpt_path checkpoints/X-Dub_model.safetensors --ref_cfg_scale 2.5 --audio_cfg_scale 10.0 --num_inference_steps 30 --output_dir ./resultsOther editing & audio projects people compare with X-Dub.
MeiGen-AI
Lip-synced video or photo from new audio, any length
KlingAIResearch
Animates a face photo or video with motion from another clip
bytedance
Re-syncs a talking-face video to another audio track
RVC-Boss
Clones a voice from a 5-second clip and reads your text in it
Do you maintain X-Dub? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/klingairesearch-x-dub)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. The paper's internal model isn't open source; this public build runs on Wan2.2-TI2V-5B and is adapted from DiffSynth-Studio. We write these descriptions ourselves. The repo's own docs are the final word.