ComfyUI Node
IndexTTS 2 / 2.5 Synthesize
IndexTTS 2/2.5 zero-shot multilingual and emotion-controllable speech synthesis.
IndexTTS 2 / 2.5 Synthesize
- model
- speaker_audio
- emotion_audio
- audio
◄text欢迎使用 IndexTTS 2.5。►
◄languageZH►
◄emotion_modesame_as_speaker►
◄emotion_weight0.65►
◄extra_emotion_text►
◄happy0.00►
◄angry0.00►
◄sad0.00►
◄fearful0.00►
◄disgusted0.00►
◄melancholic0.00►
◄surprised0.00►
◄calm0.00►
◄emotion_randomfalse►
◄duration_factor1.00►
◄interval_silence_ms200►
◄text_normalizationtrue►
◄max_text_tokens_per_segment120►
◄do_sampletrue►
◄temperature0.8►
◄top_p0.80►
◄top_k30►
◄num_beams3►
◄repetition_penalty10.0►
◄length_penalty0.0►
◄max_mel_tokens1500►
◄seed0►
Categoryaudio/IndexTTS
Inputs (30)
| Name | Type | Default | Description |
|---|---|---|---|
| model | INDEXTTS25_MODEL | — | |
| speaker_audio | AUDIO | — | |
| text | STRING | 欢迎使用 IndexTTS 2.5。 | — |
| language | COMBO | ZH | 5 options: ZH, EN, JA, AR, ES |
| emotion_mode | COMBO | same_as_speaker | 5 options: same_as_speaker, reference_audio, emotion_vector, emotion_text, extra_emotion_text |
| emotion_weight | FLOAT | 0.650–1 | — |
| extra_emotion_text | STRING | — | |
| happy | FLOAT | 0.000–1 | — |
| angry | FLOAT | 0.000–1 | — |
| sad | FLOAT | 0.000–1 | — |
| fearful | FLOAT | 0.000–1 | — |
| disgusted | FLOAT | 0.000–1 | — |
| melancholic | FLOAT | 0.000–1 | — |
| surprised | FLOAT | 0.000–1 | — |
| calm | FLOAT | 0.000–1 | — |
| emotion_random | BOOLEAN | false | — |
| duration_factor | FLOAT | 1.000.5–2 | — |
| interval_silence_ms | INT | 2000–5000 | — |
| text_normalization | BOOLEAN | true | — |
| max_text_tokens_per_segment | INT | 12020–400 | — |
| do_sample | BOOLEAN | true | — |
| temperature | FLOAT | 0.80.1–2 | — |
| top_p | FLOAT | 0.800–1 | — |
| top_k | INT | 300–100 | — |
| num_beams | INT | 31–10 | — |
| repetition_penalty | FLOAT | 10.00.1–20 | — |
| length_penalty | FLOAT | 0.0-2–2 | — |
| max_mel_tokens | INT | 150050–4096 | — |
| seed | INT | 00–9223372036854776000 | — |
| emotion_audioopt | AUDIO | — |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| audio | AUDIO | — |