Nodes/ComfyUI-NanoBanana2/NanoBanana - Text-to-Speech
ComfyUI Node

NanoBanana - Text-to-Speech

A ComfyUI node in NanoBanana2/Audio with 7 inputs and 1 output.

By IxMxAMAR·Created 5 months ago·Updated 2 months ago· 4
NanoBanana - Text-to-Speech
  • network
  • audio
api_key
modelgemini-2.5-flash-preview-tts
text
voiceKore
custom_model
style_prompt
CategoryNanoBanana2/Audio

Inputs (7)

NameTypeDefaultDescription
api_keySTRINGNanoBanana - API key. Leave blank to use GEMINI_API_KEY env var.
modelCOMBOgemini-2.5-flash-preview-ttsTTS model. Pro = higher quality, Flash = faster.
textSTRINGText to speak. Can include speaker tags for multi-speaker dialogue.
voiceCOMBOKorePrebuilt voice. Each has different characteristics.
custom_modeloptSTRING
style_promptoptSTRINGOptional style instruction prepended to text (e.g., 'Say cheerfully:').
networkoptNB_NETWORKOptional. Wire a NanoBanana - Network Route node here to route this request through that proxy (e.g. US egress).

Outputs (1)

NameTypeDescription
audioAUDIO