Nodes/ComfyUI-TBDub/TBDub Generate (video + new speech)
ComfyUI Node

TBDub Generate (video + new speech)

Redub an existing video with new speech (TBDub V1.1 Student). Returns the dubbed VIDEO (with the new speech as audio) and a JSON report.

By hiroki-abe-58·Created a day ago·Updated a day ago· 1
TBDub Generate (video + new speech)
  • runtime
  • video
  • audio
  • video
  • report
◄face_modefull_frame►
◄seed42►
◄start_frame0►
◄allow_pingpongfalse►
◄multiple_facesrefuse►
◄fps_policyresample_to_25►
CategoryTBDub

Inputs (9)

NameTypeDefaultDescription
runtimeTBDUB_RUNTIME—
videoVIDEOOne person, face visible. Its own audio is not used. Non-25 fps video is converted to 25 fps (see fps_policy).
audioAUDIOThe new speech. It becomes the output's audio track, whole and unchanged (AAC).
face_modeCOMBOfull_framefull_frame: find the face with MediaPipe (CPU), dub the 512x512 crop and paste it back into the original frames. cropped: the video is already an aligned, square face crop (it is resized to 512x512).
seedINT420–2147483647Official default 42; clip i uses seed + i.
start_frameoptINT00–100000Skip this many source frames; the speech always starts at 0.
allow_pingpongoptBOOLEANfalseIf the source is shorter than the speech, repeat it forward/backward (official behaviour). Off: refuse.
multiple_facesoptCOMBOrefuserefuse: stop if any frame shows more than one face. highest_score: official behaviour (most confident face per frame).
fps_policyoptCOMBOresample_to_25TBDub works at 25 fps. resample_to_25: convert other rates with ffmpeg's fps filter (reported). refuse: stop.

Outputs (2)

NameTypeDescription
videoVIDEO—
reportSTRING—