OpenMOSS
MOSS-TTSD
Cloned-voice dialogue audio for 1 to 5 speakers
- MOSS-TTSD v1.0
- MOSS-Audio-Tokenizer
- XY-Tokenizer
- +2
Give GPT-SoVITS a 5-second clip and it reads your text in that voice right away. About a minute of audio gets you a closer match. On a Mac, training runs on CPU for now, since Mac GPU training gave poor results.
Best for: Dubbing teams building character voices from short reference clips.
Rewritten from the README. Check the repo for the latest steps.
conda activate GPTSoVitsconda create -n GPTSoVits python=3.10bash install.sh --device <CU126|CU128|ROCM|CPU> --source <HF|HF-Mirror|ModelScope> [--download-uvr5]pwsh -F install.ps1 --Device <CU126|CU128|CPU> --Source <HF|HF-Mirror|ModelScope> [--DownloadUVR5]GPT_SoVITS/pretrained_modelspython webui.py, or double-click go-webui.bat on Windowspython GPT_SoVITS/inference_webui.pyOther editing & audio projects people compare with GPT-SoVITS.
OpenMOSS
Cloned-voice dialogue audio for 1 to 5 speakers
index-tts
Voice cloner that reads your script in five languages
hkchengrex
Makes sound effects and ambience that match your video
jianchang512
Video translated, subtitled and dubbed into another language
Do you maintain GPT-SoVITS? Add this badge to your README so English speakers can find our write-up.
[](https://openmicrodrama.com/projects/rvc-boss-gpt-sovits)The badge links to this page. Want your description changed? Email support@openmicrodrama.com.
Reviewed Oct 1, 2026. We write these descriptions ourselves. The repo's own docs are the final word.