Nodes/ComfyUI-ElevenLabs-Pro/ElevenLabs Pro - Text to Dialogue with Timestamps
ComfyUI Node

ElevenLabs Pro - Text to Dialogue with Timestamps

A ComfyUI node in ElevenLabs Pro/TTS with 32 inputs and 3 outputs.

By IxMxAMAR·Created 6 months ago·Updated 3 days ago· 1
ElevenLabs Pro - Text to Dialogue with Timestamps
    • audio
    • timestamps_json
    • voice_segments_json
    ◄api_key►
    ◄text1►
    ◄voice_id1►
    ◄modeleleven_v3►
    ◄text2►
    ◄voice_id2►
    ◄text3►
    ◄voice_id3►
    ◄text4►
    ◄voice_id4►
    ◄text5►
    ◄voice_id5►
    ◄text6►
    ◄voice_id6►
    ◄text7►
    ◄voice_id7►
    ◄text8►
    ◄voice_id8►
    ◄text9►
    ◄voice_id9►
    ◄text10►
    ◄voice_id10►
    ◄stability0.50►
    ◄apply_text_normalizationoff►
    ◄languageAuto Detect►
    ◄output_formatmp3_44100_192►
    ◄seed0►
    ◄enable_loggingtrue►
    ◄previous_text►
    ◄future_text►
    ◄use_pvc_as_ivcfalse►
    ◄pronunciation_dictionary_locators►
    CategoryElevenLabs Pro/TTS

    Inputs (32)

    NameTypeDefaultDescription
    api_keySTRING—
    text1STRINGSpeaker 1 text.
    voice_id1STRINGSpeaker 1 voice ID.
    modelCOMBOeleven_v32 options: eleven_v3, eleven_v4
    text2optSTRING—
    voice_id2optSTRING—
    text3optSTRING—
    voice_id3optSTRING—
    text4optSTRING—
    voice_id4optSTRING—
    text5optSTRING—
    voice_id5optSTRING—
    text6optSTRING—
    voice_id6optSTRING—
    text7optSTRING—
    voice_id7optSTRING—
    text8optSTRING—
    voice_id8optSTRING—
    text9optSTRING—
    voice_id9optSTRING—
    text10optSTRING—
    voice_id10optSTRING—
    stabilityoptFLOAT0.500–1Voice stability. Lower = more expressive/emotional, Higher = more consistent/monotone. Creative(<0.5), Natural(0.5), Robust(>0.5).
    apply_text_normalizationoptCOMBOoffeleven_v3 requires 'off' — keeping default.
    languageoptCOMBOAuto Detect33 options: Auto Detect, English (en), Arabic (ar), Bulgarian (bg), Chinese (zh), Croatian (hr), +27
    output_formatoptCOMBOmp3_44100_192Audio output format. mp3_44100_192 and opus require Creator tier+.
    seedoptINT00–4294967295Seed for reproducibility. 0 = random. Determinism not guaranteed.
    enable_loggingoptBOOLEANtrueIf False, requests zero-retention mode (audio + text not stored by ElevenLabs). Required for HIPAA / privacy-sensitive content.
    previous_textoptSTRINGContext only — up to 100 characters that come right BEFORE this dialogue, for continuity. Not supported by every model.
    future_textoptSTRINGContext only — up to 100 characters that come right AFTER this dialogue, for continuity. Not supported by every model.
    use_pvc_as_ivcoptBOOLEANfalseUse the IVC version of a Professional Voice Clone.
    pronunciation_dictionary_locatorsoptSTRINGJSON array of {"pronunciation_dictionary_id": ..., "version_id": ...} objects (up to 3).

    Outputs (3)

    NameTypeDescription
    audioAUDIO—
    timestamps_jsonSTRING—
    voice_segments_jsonSTRING—