DIGIT ElevenLabs Dialogue
Two voices, one node — multi-speaker dialogue straight from a script
- audio
Here's the problem the DIGIT ElevenLabs Dialogue node solves: you need two characters talking to each other, and the naive approach is two text-to-speech nodes, two separate runs, and a pile of manual alignment. This node does it in one shot - you give it up to ten dialogue lines, each assigned to a voice ID, and it returns a single combined audio track. No stitching, no timing math, no "did the pause land right" anxiety.
It's the multi-speaker member of the pack's ElevenLabs family, and the payoff is most obvious in the obvious place: a scripted conversation - a product ad, a character scene, a two-person narration. The voices are ElevenLabs voices (the commercial bar for this kind of work), so the quality ceiling is high, and the node handles the layout so you don't have to.
How it works
The inputs mirror a script: text1 with its voice_id1, and so on up to text10/voice_id10. num_entries (default 2, max 10) tells the node how many lines to actually use, so you can leave the rest of the text fields empty and just flip the count. Each line goes to ElevenLabs with its assigned voice, and the segments come back assembled into one audio output (a ComfyUI AUDIO tensor, PCM 44.1kHz by default).
The shared voice controls: stability (0.5 default - lower for more expressiveness, higher for consistency), seed for reproducible takes, and model (eleven_v3). There's also language_code for non-English dialogue and apply_text_normalization (auto/on/off) for how numbers and abbreviations are expanded. output_format lets you switch between pcm_44100, mp3_44100_192, and opus_48000_192 - the MP3/Opus options are worth it if you're sending the audio straight to a video encoder.
The api_key field is optional because the node auto-detects ELEVENLABS_API_KEY (or the pack's DIGIT_ELEVENLABS_API_KEY) from the environment - paste it on the node only if you're not using env vars.
Installing it
Standard pack install:
cd ComfyUI/custom_nodes
git clone https://github.com/thedepartmentofexternalservices/comfyui-digit.git
cd comfyui-digit
pip install -r requirements.txt
Or ComfyUI Manager → search comfyui-digit → install → restart. Then set your key before starting ComfyUI:
export ELEVENLABS_API_KEY=your_key_here
Where to get the voice IDs
The natural pairing is the pack's DIGIT ElevenLabs Voice Selector node, which outputs a voice_id you can wire straight into voice_id1, voice_id2, etc. - or paste IDs you already have from the ElevenLabs site. That's the whole trick to this node: the per-line voice assignment is what makes a two-voice track sound like two different people instead of one person doing voices. Set the lines, assign the voices, hit run, and the single audio output is ready for your video or your save node.
Inputs (28)
| Name | Type | Default | Description |
|---|---|---|---|
| text1 | STRING | Dialogue line 1. | |
| voice_id1 | STRING | Voice ID for line 1. | |
| num_entries | INT | 21–10 | Number of dialogue entries to use. |
| model | COMBO | eleven_v3 | 1 options: eleven_v3 |
| stability | FLOAT | 0.500–1 | — |
| seed | INT | 10–4294967295 | — |
| api_keyopt | STRING | ElevenLabs API key. Auto-detected from ELEVENLABS_API_KEY env var. | |
| text2opt | STRING | — | |
| voice_id2opt | STRING | — | |
| text3opt | STRING | — | |
| voice_id3opt | STRING | — | |
| text4opt | STRING | — | |
| voice_id4opt | STRING | — | |
| text5opt | STRING | — | |
| voice_id5opt | STRING | — | |
| text6opt | STRING | — | |
| voice_id6opt | STRING | — | |
| text7opt | STRING | — | |
| voice_id7opt | STRING | — | |
| text8opt | STRING | — | |
| voice_id8opt | STRING | — | |
| text9opt | STRING | — | |
| voice_id9opt | STRING | — | |
| text10opt | STRING | — | |
| voice_id10opt | STRING | — | |
| language_codeopt | STRING | — | |
| apply_text_normalizationopt | COMBO | auto | 3 options: auto, on, off |
| output_formatopt | COMBO | pcm_44100 | 3 options: pcm_44100, mp3_44100_192, opus_48000_192 |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| audio | AUDIO | — |