ComfyUI Node
⭐ Star LTXV All-in-One (2-Pass)
All-in-one LTXV two-pass sampler (Sulphur workflow port). Pass 1 at half resolution -> 2x latent upscale -> pass 2 at full resolution. T2V / I2V / I2V+Audio, HD/FHD presets with ratio-from-image, model+LoRA+CLIP+VAE caching built in.
⭐ Star LTXV All-in-One (2-Pass)
- image
- audio
- model_override
- images
- audio
- frame_rate
◄mode▶️ image_to_video►
◄positive_prompt►
◄negative_promptconsole game, video game, cartoon, childish, ugly►
◄base_model▾►
◄clip_1▾►
◄clip_2▾►
◄vae▾►
◄audio_vae▾►
◄upscale_model▾►
◄video_sizeHD►
◄ratio1:1►
◄ratio_from_imagetrue►
◄custom_width1024►
◄custom_height1024►
◄frame_rate25►
◄seconds10►
◄seed0►
◄sigma_preset▾►
◄override_audiofalse►
◄lora_1▾►
◄lora_1_strength0.60►
◄lora_2▾►
◄lora_2_strength1.00►
◄lora_3▾►
◄lora_3_strength1.00►
◄custom_sigmas_pass11.0, 0.995833, 0.991667, 0.9875, 0.983333, 0.979167, 0.975, 0.93125, 0.847917, 0.725, 0.522917, 0.28125, 0.0►
◄sigmas_pass20.85, 0.725, 0.6, 0.4219, 0.0►
◄cfg1.0►
◄sampler_pass1euler_ancestral_cfg_pp►
◄sampler_pass2euler_cfg_pp►
◄weight_dtype▾►
Category⭐StarNodes/Video
Inputs (34)
| Name | Type | Default | Description |
|---|---|---|---|
| mode | COMBO | ▶️ image_to_video | text_to_video: prompt only. image_to_video: connect an image. image_audio_to_video: connect an image AND an audio file. |
| positive_prompt | STRING | What you want to see. LTXV likes detailed, film-style descriptions with timestamps. | |
| negative_prompt | STRING | console game, video game, cartoon, childish, ugly | What to avoid. Default is the negative prompt from the original workflow. |
| base_model | COMBO | LTXV 2.3 A/V checkpoint from models/diffusion_models (e.g. sulphur2Base). Reloaded only when the selection changes. | |
| clip_1 | COMBO | Main text encoder from models/text_encoders (e.g. gemma-3-12b ... int4). | |
| clip_2 | COMBO | LTXV text projection from models/text_encoders (e.g. ltx-2.3_text_projection). | |
| vae | COMBO | Video VAE from models/vae (e.g. LTX23_video_vae). | |
| audio_vae | COMBO | Audio VAE from models/vae (e.g. LTX23_audio_vae). | |
| upscale_model | COMBO | Latent upscaler from models/latent_upscale_models (e.g. ltx-2.3-spatial-upscaler-x2). Used between the passes. | |
| video_size | COMBO | HD | HD ~1280px, FHD ~1920px (same tables as the Star LTX Video Settings node), Custom = custom_width/height below. |
| ratio | COMBO | 1:1 | Aspect ratio. Overridden by the input image's ratio when 'ratio_from_image' is enabled and an image is connected. |
| ratio_from_image | BOOLEAN | true | Pick the closest preset ratio to the connected image. Falls back to 'ratio' when no image is connected. |
| custom_width | INT | 102432–8192 | Only used when video_size = Custom. |
| custom_height | INT | 102432–8192 | Only used when video_size = Custom. |
| frame_rate | INT | 251–120 | Frames per second of the output video. |
| seconds | INT | 101–120 | Video length in seconds. Frame count is snapped to 8n+1 (4s @ 25fps = 97 frames). |
| seed | INT | 00–18446744073709550000 | Shared by both sampling passes. |
| sigma_preset | COMBO | First-pass noise schedule - the three presets from the original workflow's note node. 12 = default, 8 = faster, 16 = finer. 'custom' uses custom_sigmas_pass1 below. | |
| imageopt | IMAGE | Start frame / guide image (image_to_video modes). | |
| audioopt | AUDIO | Voice / music track (image_audio_to_video mode). Trimmed to the video length and preserved as-is. | |
| override_audioopt | BOOLEAN | false | text_to_video / image_to_video only: when disabled (default), the connected 'audio' input is ignored and the model-generated audio is sent to the audio output. When enabled, the connected 'audio' input is passed straight to the audio output instead. Ignored in image_audio_to_video mode, where the connected audio is always passed through to the output. |
| lora_1opt | COMBO | Optional LoRA stack, applied in order 1 -> 3. | |
| lora_1_strengthopt | FLOAT | 0.60-100–100 | The distilled LoRA in the original workflow ran at 0.6. |
| lora_2opt | COMBO | 1 options: None | |
| lora_2_strengthopt | FLOAT | 1.00-100–100 | — |
| lora_3opt | COMBO | 1 options: None | |
| lora_3_strengthopt | FLOAT | 1.00-100–100 | — |
| custom_sigmas_pass1opt | STRING | 1.0, 0.995833, 0.991667, 0.9875, 0.983333, 0.979167, 0.975, 0.93125, 0.847917, 0.725, 0.522917, 0.28125, 0.0 | Only used when sigma_preset = custom. |
| sigmas_pass2opt | STRING | 0.85, 0.725, 0.6, 0.4219, 0.0 | Second-pass (refine) schedule. Default from the workflow. |
| cfgopt | FLOAT | 1.00–100 | Both passes. 1.0 for distilled models, as in the workflow. |
| sampler_pass1opt | COMBO | euler_ancestral_cfg_pp | Sampler for pass 1 (half resolution). |
| sampler_pass2opt | COMBO | euler_cfg_pp | Sampler for pass 2 (full resolution refine). |
| weight_dtypeopt | COMBO | Override base-model dtype. 'default' = as stored. | |
| model_overrideopt | MODEL | Optional external model (e.g. patched with flash/sage attention). When connected, this is used instead of loading 'base_model' from the dropdown, and the LoRA stack below is applied to it directly. |
Outputs (3)
| Name | Type | Description |
|---|---|---|
| images | IMAGE | — |
| audio | AUDIO | — |
| frame_rate | FLOAT | — |