Nodes/Artha-Gemini/πŸ”± Gemini Speech
ComfyUI Node

πŸ”± Gemini Speech

Generates speech from text using the Gemini TTS model.

By CyrostarΒ·Created about a year agoΒ·Updated 12 months agoΒ· 1
πŸ”± Gemini Speech
    • audio
    β—„text_promptA cat with a hatβ–Ί
    β—„voiceKoreβ–Ί
    β—„api_keyβ–Ί
    β—„modelgemini-2.5-flash-preview-ttsβ–Ί
    β—„max_tokens5000β–Ί
    β—„temperature0.7β–Ί
    CategoryArtha/LLM/GEMINI

    Inputs (6)

    NameTypeDefaultDescription
    text_promptSTRINGA cat with a hatβ€”
    voiceCOMBOKore30 options: Zephyr, Autonoe, Puck, Laomedeia, Charon, Rasalgethi, +24
    api_keySTRINGAPI key will be visible in plain text. Consider adding your api to the api.json located inside this custom node folder.
    modelCOMBOgemini-2.5-flash-preview-tts1 options: gemini-2.5-flash-preview-tts
    max_tokensINT50001–8192For Gemini models, a token is equivalent to about 4 characters. 100 tokens is equal to about 60-80 English words.
    temperatureFLOAT0.70–2A temperature of 0 means only the most likely tokens are selected, and there's no randomness. Conversely, a high temperature injects a high degree of randomness into the tokens selected by the model, leading to more unexpected, surprising model responses.

    Outputs (1)

    NameTypeDescription
    audioAUDIOβ€”