OpenMOSS
MOSS-TTSD
Cloned-voice dialogue audio for 1 to 5 speakers
- MOSS-TTSD v1.0
- MOSS-Audio-Tokenizer
- XY-Tokenizer
- +2
Give GPT-SoVITS a 5-second clip and it reads your text in that voice right away. About a minute of audio gets you a closer match. On a Mac, training runs on CPU for now, since Mac GPU training gave poor results.
Best for: Dubbing teams building character voices from short reference clips.
Rewritten from the README. Check the repo for the latest steps.
conda activate GPTSoVitsconda create -n GPTSoVits python=3.10bash install.sh --device <CU126|CU128|ROCM|CPU> --source <HF|HF-Mirror|ModelScope> [--download-uvr5]pwsh -F install.ps1 --Device <CU126|CU128|CPU> --Source <HF|HF-Mirror|ModelScope> [--DownloadUVR5]GPT_SoVITS/pretrained_modelspython webui.py, or double-click go-webui.bat on Windowspython GPT_SoVITS/inference_webui.pyOther editing & audio projects people compare with GPT-SoVITS.
OpenMOSS
Cloned-voice dialogue audio for 1 to 5 speakers
index-tts
Voice cloner that reads your script in five languages
hkchengrex
Makes sound effects and ambience that match your video
k4yt3x
Makes small video bigger and smoother using Anime4K, Real-ESRGAN, RIFE
Do you maintain GPT-SoVITS? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/rvc-boss-gpt-sovits)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. We write these descriptions ourselves. The repo's own docs are the final word.