ComfyUI Node
ποΈ ThinkSound Sampler
A ComfyUI node in ThinkSound with 11 inputs and 1 output.
ποΈ ThinkSound Sampler
- thinksound_model
- feature_utils
- video
- audio
βduration8.0βΊ
βsteps24βΊ
βcfg_scale5.0βΊ
βseed0βΊ
βcaptionβΊ
βcot_descriptionβΊ
βforce_offloadtrueβΊ
βperformance_modebalancedβΊ
CategoryThinkSound
Inputs (11)
| Name | Type | Default | Description |
|---|---|---|---|
| thinksound_model | THINKSOUND_MODEL | β | |
| feature_utils | THINKSOUND_FEATUREUTILS | β | |
| duration | FLOAT | 8.01β30 | Duration of generated audio in seconds |
| steps | INT | 241β100 | Number of denoising steps (more = better quality, slower) |
| cfg_scale | FLOAT | 5.01β20 | Classifier-free guidance scale (higher = more faithful to text) |
| seed | INT | 00β18446744073709550000 | Random seed for reproducible results |
| caption | STRING | Short description of desired audio (e.g., 'dog barking', 'ocean waves') | |
| cot_description | STRING | Detailed chain-of-thought description for enhanced audio generation | |
| force_offload | BOOLEAN | true | Offload models after generation to save VRAM |
| performance_mode | COMBO | balanced | Generation performance profile |
| videoopt | IMAGE | Input video frames for video-to-audio generation (optional) |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| audio | AUDIO | β |