modelscope
DiffSynth-Studio
Trains and runs image, video and audio diffusion models on your own GPU
- MiniMax H3
- LTX-2.5
- Wan-Animate-2
- +7
VideoX-Fun makes AI images and videos, and trains your own video models. It runs text-to-video, image-to-video, video-to-video, and control videos from Canny, Pose or Depth. You drive it with Python scripts, a Gradio web UI, or ComfyUI nodes. Weights take about 60GB of disk, and big models need offload or many GPUs.
Best for: Builders and researchers who want to run or fine-tune open video models on their own GPUs.
Rewritten from the README. Check the repo for the latest steps.
--gpus all.docker pull mybigpai-public-registry.cn-beijing.cr.aliyuncs.com/easycv/torch_cuda:cogvideox_funcd VideoX-Fun.git clone https://github.com/aigc-apps/VideoX-Fun.gitmodels/Diffusion_Transformer and models/Personalized_Model, then download weights from Hugging Face or ModelScope into them.examples/cogvideox_fun/predict_t2v.py, run it, and find videos in samples/cogvideox-fun-videos.examples/cogvideox_fun/app.py to open the Gradio interface and generate there.xfuser==0.4.2 and yunchang==0.6.2, then run the command below.torchrun --nproc-per-node=8 examples/wan2.1_fun/predict_t2v.pyOther model tooling projects people compare with VideoX-Fun.
modelscope
Trains and runs image, video and audio diffusion models on your own GPU
ModelTC
Loads open image and video models onto your own GPU
Lightricks
Local LTX video generation and editing on your desktop
Wan-Video
Open video models for text, image, speech and character animation
Do you maintain VideoX-Fun? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/aigc-apps-videox-fun)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. We write these descriptions ourselves. The repo's own docs are the final word.