WeChatCV
Stand-In
Wan add-on that keeps one face consistent across shots
- Wan 2.1 T2V 14B
- Wan 2.2 T2V A14B
- VACE
- +2
ShotStream is research code that generates multi-shot video frames as you go. It's built on Wan2.1-T2V-1.3B and reports 16 FPS on one NVIDIA GPU. You get inference and training scripts, but it's a reference implementation, not an app. Training needs multi-node GPUs and edits to the bash scripts.
Best for: Researchers and engineers who want to run, fine-tune or study a streaming multi-shot video model.
Rewritten from the README. Check the repo for the latest steps.
conda activate shotstreamgit clone https://github.com/KlingAIResearch/ShotStream.gitconda create -n shotstream python=3.10 -ybash tools/setup/env.shpip install -r requirements.txtpip install flash-attn --no-build-isolationwan_models and ckpts with git-lfs. Or run the command belowbash tools/setup/download_ckpt.shbash tools/inference/causal_fewsteps.shMASTER_ADDR in the bash scripts, then start with the command belowbash tools/train/1_basemodel.sh 0Other model tooling projects people compare with ShotStream.
WeChatCV
Wan add-on that keeps one face consistent across shots
Kevin-thu
Give it a shot list, get a minute-long multi-shot video
modelscope
Trains and runs image, video and audio diffusion models on your own GPU
Wan-Video
Open video models for text, image, speech and character animation
Do you maintain ShotStream? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/klingairesearch-shotstream)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. Reference implementation of an ECCV 2026 paper from MMLab CUHK and the Kling team; built on Wan2.1-T2V-1.3B. We write these descriptions ourselves. The repo's own docs are the final word.