Skip to content

index-tts /index-tts

Voice cloner that reads your script in five languages
Editing & audioCustom licenseEnglish README
Stars
24k
30 days
+8
Last push
3 days ago
index-tts screenshot from its README
From the index-tts README on GitHub.

About index-tts

IndexTTS copies a voice from one short clip, then reads your script in it. Version 2.5 covers Chinese, English, Japanese, Spanish and Arabic, and you can steer emotion, speed and hard-word pronunciation. An NVIDIA GPU with CUDA 12.8 or newer is required. The license is custom, so check it before commercial use.

Best for: Short-drama and comic-drama dubbers who need cloned voices in five languages.

What it does

  • Clones a voice from one reference clip, with no training per speaker
  • Sets emotion with a sad or happy reference clip, an 8-value vector, or text
  • emo_alpha (0.0-1.0) controls how strongly the emotion reference pulls the voice
  • Changes speaking speed with duration_factor, from 0.5x to 2.0x
  • Fixes pronunciation with Pinyin, CMU phonemes or Japanese Kana
  • Runs as a Gradio WebUI, a Python API, or a vLLM server

Quickstart

Rewritten from the README. Check the repo for the latest steps.

  1. 1
    Clone the repo: the command below
    git clone https://github.com/index-tts/index-tts.git && cd index-tts
  2. 2
    Install uv, then sync dependencies: pip install -U uv and uv sync --all-extras
  3. 3
    Download the weights: the command below
    hf download IndexTeam/IndexTTS-2.5 --local-dir=checkpoints
  4. 4
    Check your GPU with uv run tools/gpu_check.py
  5. 5
    Start the WebUI with uv run webui.py, then open http://127.0.0.1:7860
  6. 6
    Or run the CLI: the command below
    PYTHONPATH="$PYTHONPATH:." uv run indextts/infer_v2_5.py --cfg_path checkpoints/config.yaml --model_dir checkpoints --text "Hello world" --lang EN

Models and languages

Models it works with

  • IndexTTS-2.5
  • IndexTTS-2
  • IndexTTS-1.5
  • IndexTTS-1.0
  • Qwen

Interface and docs

  • Chinese
  • English
  • Japanese
  • Spanish
  • Arabic

Head-to-head

Alternatives

Other editing & audio projects people compare with index-tts.

OpenMOSS

MOSS-TTSD

Cloned-voice dialogue audio for 1 to 5 speakers

Editing & audio
  • MOSS-TTSD v1.0
  • MOSS-Audio-Tokenizer
  • XY-Tokenizer
  • +2
1.4k GitHub stars26 days agoApache-2.0

xcLee001

SonicVale

Dubs novels and scripts with multi-character AI voices

Editing & audio
  • IndexTTS-2
628 GitHub stars1 month agoAGPL-3.0

RVC-Boss

GPT-SoVITS

Clones a voice from a 5-second clip and reads your text in it

Editing & audio
  • GPT-SoVITS v2Pro / v2ProPlus
  • GPT-SoVITS v4
  • GPT-SoVITS v3
  • +7
62k GitHub stars+16/30d1 month agoMIT

jianchang512

pyvideotrans

Video translated, subtitled and dubbed into another language

Editing & audio
  • Faster-Whisper
  • WhisperX
  • Parakeet
  • +7
19k GitHub stars+6/30d2 days agoGPL-3.0

Featured on OpenMicroDrama

Do you maintain index-tts? Add this badge to your README so English speakers can find our write-up.

Featured on OpenMicroDrama
markdown
[![Featured on OpenMicroDrama](https://openmicrodrama.com/badges/featured.svg)](https://openmicrodrama.com/projects/index-tts-index-tts)

The badge links to this page. Want your description changed? Email support@openmicrodrama.com.

Reviewed Oct 1, 2026. Custom bilibili Model Use License. Check its terms before commercial use. We write these descriptions ourselves. The repo's own docs are the final word.