ComfyUI Node
NanoBanana - Text-to-Speech
A ComfyUI node in NanoBanana2/Audio with 7 inputs and 1 output.
NanoBanana - Text-to-Speech
- network
- audio
◄api_key►
◄modelgemini-2.5-flash-preview-tts►
◄text►
◄voiceKore►
◄custom_model►
◄style_prompt►
CategoryNanoBanana2/Audio
Inputs (7)
| Name | Type | Default | Description |
|---|---|---|---|
| api_key | STRING | NanoBanana - API key. Leave blank to use GEMINI_API_KEY env var. | |
| model | COMBO | gemini-2.5-flash-preview-tts | TTS model. Pro = higher quality, Flash = faster. |
| text | STRING | Text to speak. Can include speaker tags for multi-speaker dialogue. | |
| voice | COMBO | Kore | Prebuilt voice. Each has different characteristics. |
| custom_modelopt | STRING | — | |
| style_promptopt | STRING | Optional style instruction prepended to text (e.g., 'Say cheerfully:'). | |
| networkopt | NB_NETWORK | Optional. Wire a NanoBanana - Network Route node here to route this request through that proxy (e.g. US egress). |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| audio | AUDIO | — |