ComfyUI Node
Zonos Generate
A ComfyUI node in audio with 16 inputs and 1 output.
Zonos Generate
- sample_audio
- prefix_audio
- emotion
- AUDIO
◄speechThis is what I want to say►
◄seed1►
◄model_type▾►
◄language▾►
◄pitch_std20►
◄speaking_rate15►
◄dnsmos_ovrl4.0►
◄cfg_scale2.0►
◄min_p0.15►
◄speed1.00►
◄disable_compilertrue►
◄sample_textText of sample_audio►
◄speaker_noisedfalse►
Categoryaudio
Inputs (16)
| Name | Type | Default | Description |
|---|---|---|---|
| speech | STRING | This is what I want to say | — |
| seed | INT | 1 | Seed. -1 = random |
| model_type | COMBO | 2 options: Zyphra/Zonos-v0.1-transformer, Zyphra/Zonos-v0.1-hybrid | |
| language | COMBO | 109 options: af, am, an, ar, as, az, +103 | |
| pitch_std | FLOAT | 200–300 | — |
| speaking_rate | FLOAT | 155–30 | — |
| dnsmos_ovrl | FLOAT | 4.01–5 | — |
| cfg_scale | FLOAT | 2.01–5 | — |
| min_p | FLOAT | 0.150–1 | — |
| speed | FLOAT | 1.00 | Speed. >1.0 slower. <1.0 faster |
| disable_compiler | BOOLEAN | true | Disable PyTorch compiler for better compatibility |
| sample_audioopt | AUDIO | — | |
| sample_textopt | STRING | Text of sample_audio | — |
| prefix_audioopt | AUDIO | Optional audio to continue from | |
| speaker_noisedopt | BOOLEAN | false | Apply denoising to speaker reference |
| emotionopt | EMOTION | — |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| AUDIO | AUDIO | — |