WeChatCV
Stand-In
Wan add-on that keeps one face consistent across shots
- Wan 2.1 T2V 14B
- Wan 2.2 T2V A14B
- VACE
- +2

ShotStream is research code that generates multi-shot video frames as you go. It's built on Wan2.1-T2V-1.3B and reports 16 FPS on one NVIDIA GPU. You get inference and training scripts, but it's a reference implementation, not an app. Training needs multi-node GPUs and edits to the bash scripts.
Best for: Researchers and engineers who want to run, fine-tune or study a streaming multi-shot video model.
Rewritten from the README. Check the repo for the latest steps.
conda activate shotstreamgit clone https://github.com/KlingAIResearch/ShotStream.gitconda create -n shotstream python=3.10 -ybash tools/setup/env.shpip install -r requirements.txtpip install flash-attn --no-build-isolationwan_models and ckpts with git-lfs. Or run the command belowbash tools/setup/download_ckpt.shbash tools/inference/causal_fewsteps.shMASTER_ADDR in the bash scripts, then start with the command belowbash tools/train/1_basemodel.sh 0Other model tooling projects people compare with ShotStream.
WeChatCV
Wan add-on that keeps one face consistent across shots
Kevin-thu
Give it a shot list, get a minute-long multi-shot video
modelscope
Trains and runs image, video and audio diffusion models on your own GPU
deepbeepmeep
Desktop app that runs open video, image and audio models on your own PC
Do you maintain ShotStream? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/klingairesearch-shotstream)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. Reference implementation of an ECCV 2026 paper from MMLab CUHK and the Kling team; built on Wan2.1-T2V-1.3B. We write these descriptions ourselves. The repo's own docs are the final word.