Nodes/ComfyUI-ThinkSound_Wrapper/πŸŽ›οΈ ThinkSound Sampler
ComfyUI Node

πŸŽ›οΈ ThinkSound Sampler

A ComfyUI node in ThinkSound with 11 inputs and 1 output.

By ShmuelRonenΒ·Created about a year agoΒ·Updated about a year agoΒ· 22
πŸŽ›οΈ ThinkSound Sampler
  • thinksound_model
  • feature_utils
  • video
  • audio
β—„duration8.0β–Ί
β—„steps24β–Ί
β—„cfg_scale5.0β–Ί
β—„seed0β–Ί
β—„captionβ–Ί
β—„cot_descriptionβ–Ί
β—„force_offloadtrueβ–Ί
β—„performance_modebalancedβ–Ί
CategoryThinkSound

Inputs (11)

NameTypeDefaultDescription
thinksound_modelTHINKSOUND_MODELβ€”
feature_utilsTHINKSOUND_FEATUREUTILSβ€”
durationFLOAT8.01–30Duration of generated audio in seconds
stepsINT241–100Number of denoising steps (more = better quality, slower)
cfg_scaleFLOAT5.01–20Classifier-free guidance scale (higher = more faithful to text)
seedINT00–18446744073709550000Random seed for reproducible results
captionSTRINGShort description of desired audio (e.g., 'dog barking', 'ocean waves')
cot_descriptionSTRINGDetailed chain-of-thought description for enhanced audio generation
force_offloadBOOLEANtrueOffload models after generation to save VRAM
performance_modeCOMBObalancedGeneration performance profile
videooptIMAGEInput video frames for video-to-audio generation (optional)

Outputs (1)

NameTypeDescription
audioAUDIOβ€”