Extensions/ComfyUI-LTXVideo
ComfyUI Extension Runs on cloud

ComfyUI-LTXVideo

Custom nodes for LTX-Video support in ComfyUI

By LightricksΒ·Created 2 years agoΒ·Updated about a month agoΒ· 3,956
Lightricks/ComfyUI-LTXVideo
Nodes78
On cloudRunnable
Categorymodel/conditioning/ltxv, lightricks/LTXV
Stars3,956
Updatedabout a month ago

Nodes (78)

LTXVAddGuide

Pin a keyframe into your LTX video

model/conditioning/ltxv
πŸ…›πŸ…£πŸ…§ APG Guider

Push guidance on LTX without the blown-out look

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ Dynamic Conditioning

Dial LTX's conditioning strength to fight I2V drift

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ Gemma API Text Encode

Skip the 22GB text encoder

api node/text/Lightricks
πŸ…›πŸ…£πŸ…§ Guider Parameters

The config brick for Multimodal Guider

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ Image to CPU

Move an image off the GPU to free VRAM in tight LTX workflows

utility
πŸ…›πŸ…£πŸ…§ Linear transition with overlap

Crossfade video latents to stitch clips without a seam

Lightricks/latent
πŸ…›πŸ…£πŸ…§ Low VRAM Audio VAE Loader

Load the audio decoder without blowing your budget

LTXV/loaders
πŸ…›πŸ…£πŸ…§ Low VRAM Checkpoint Loader

Fit LTX in 32GB

LTXV/loaders
Low VRAMLoad Latent Upscale Model

Load the upscaler without OOM

LTXV/loaders
πŸ…›πŸ…£πŸ…§ Add Video IC-LoRA Guide

Feed a control video into LTX

Lightricks/IC-LoRA
πŸ…›πŸ…£πŸ…§ Add Video IC-LoRA Guide Advanced

Per-guide attention control

Lightricks/IC-LoRA
LTX Attention Bank

Stash attention during inversion to keep edits consistent

ltxtricks
LTX Attn Block Override

Pick which LTX blocks get attention injection

ltxtricks
LTX Attention Override

Pick which attention layers LTX's tricks act on

ltxtricks/attn
LTX Feta Enhance

More detail for basically free

ltxtricks
πŸ…›πŸ…£πŸ…§ Float To Int

The tiny converter that keeps LTX graphs valid

math/conversion
LTX Flow Edit CFG Guider

Edit a video by swapping the prompt, no inversion

ltxtricks
LTX Flow Edit Sampler

Edit a video toward a new prompt without full inversion

ltxtricks
LTX Forward Model Pred

The model flip that makes video editing work

ltxtricks
πŸ…›πŸ…£πŸ…§ IC-LoRA Loader Model Only

Load an LTX control LoRA

Lightricks/IC-LoRA
LTX Apply Perturbed Attention

Cleaner structure without a negative prompt

ltxtricks/attn
LTX Prepare Attn Injection

Keep the source structure when you FlowEdit an LTX video

fluxtapoz
πŸ…›πŸ…£πŸ…§ LTXQ8Patch

Apply LTX's Q8 quantized kernels for speed and lower VRAM

lightricks/LTXV
LTX Reverse Model Pred

The inversion pass behind LTX FlowEdit

ltxtricks
LTX Rf-Inv Forward Sampler

Invert a real clip into noise for LTX video editing

ltxtricks
LTX Rf-Inv Reverse Sampler

Regenerate an inverted LTX clip toward a new prompt

ltxtricks
πŸ…›πŸ…£πŸ…§ LTXV Adain Latent

Color-match LTX latents to fix drift between stages and clips

Lightricks/latents
πŸ…›πŸ…£πŸ…§ LTXV Add Guide Advanced

Keyframe conditioning with preprocessing

conditioning/video_models
πŸ…›πŸ…£πŸ…§ LTXV Add Guide Advanced Attention

Pin a frame, control how hard it sticks

conditioning/video_models
πŸ…›πŸ…£πŸ…§ LTXV Add Latent Guide

Anchor a video with an encoded latent

ltxtricks
πŸ…›πŸ…£πŸ…§ LTXV Add Latents

Concatenate two LTX video latents into one longer clip

latent/video
πŸ…›πŸ…£πŸ…§ LTXV Apply STG

Pick which layers spatio-temporal guidance touches

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXV Audio Only Empty Video Latent

The throwaway video latent LTX-2 needs for text-to-audio

Lightricks/audio
πŸ…›πŸ…£πŸ…§ LTXV Audio Only Model

Turn LTX-2 into a text-to-audio generator

Lightricks/audio
πŸ…›πŸ…£πŸ…§ LTXV Base Sampler

The all-in-one LTX video sampler

sampling
πŸ…›πŸ…£πŸ…§ LTXV Dilate Latent

Stretch an LTX latent on a grid

latent/video
πŸ…›πŸ…£πŸ…§ LTXV Dilate Video Mask

Grow a mask in space and across frames

Lightricks/mask_operations
πŸ…›πŸ…£πŸ…§ LTX Draw Sparse Tracks

Turn point trajectories into a motion-track control video for LTX

Lightricks/motion_tracking
πŸ…›πŸ…£πŸ…§ LTXV Extend Sampler

Stitch a longer clip out of a short one

sampling
πŸ…›πŸ…£πŸ…§ Gemma 3 Model Loader

Load the LTX-2 text encoder

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ Gemma 3 Prompt Enhancer

Auto-expand LTX-2 prompts

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXVHDR Decode Postprocess

Turn LTX-2.3's HDR IC-LoRA output into real EXR frames

Lightricks/HDR
πŸ…›πŸ…£πŸ…§ LTXV Img To Video Advanced

The I2V node with the knobs that actually matter

conditioning/video_models
πŸ…›πŸ…£πŸ…§ LTXV Img To Video Condition Only

Pin a video's opening frames to your image

conditioning/video_models
πŸ…›πŸ…£πŸ…§ LTXV In Context Sampler

The sampler that drives LTX-2 IC-LoRA control

sampling
πŸ…›πŸ…£πŸ…§ LTXV Inpaint Preprocess

Green-screen the region you want regenerated

Lightricks/image_processing
πŸ…›πŸ…£πŸ…§ LTX Laplacian Pyramid Blend

Seamless masked image blends with no visible seam

Lightricks/utility
πŸ…›πŸ…£πŸ…§ LTXV Linear Overlap Latent Transition

Crossfade two video latents for smooth long clips

Lightricks/latent
πŸ…›πŸ…£πŸ…§ LTXV Load Conditioning

Reuse an encoding instead of re-running Gemma

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXV Looping Sampler

Long and looping LTX clips

sampling
πŸ…›πŸ…£πŸ…§ LTXV Multi Prompt Provider

Encode several prompts at once for multi-segment LTX video

prompt
πŸ…›πŸ…£πŸ…§ LTXV Normalizing Sampler

Keep LTX-2 audio and video latents balanced while sampling

utility
πŸ…›πŸ…£πŸ…§ LTXV Patcher VAE

Wire your VAE through it so LTX decode behaves

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXV Per Step Adain Patcher

Anchor LTX color and tone to a reference, step by step

Lightricks/latents
πŸ…›πŸ…£πŸ…§ LTXV Per Step Stat Norm Patcher

Stop LTX latents from blowing out mid-sample

Lightricks/latents
πŸ…›πŸ…£πŸ…§ LTXV Preprocess Masks

Get your masks into LTX's latent space

Lightricks/mask_operations
πŸ…›πŸ…£πŸ…§ LTXV Prompt Enhancer

Auto-expand your LTX prompt

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXV Prompt Enhancer Loader

Turn a short prompt into the paragraph LTX wants

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXVQ8Lora Model Loader

Load an LTX LoRA in Q8 without the kernel headache

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXV Save Conditioning

Encode the prompt once, reuse it forever

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ LTXV Select Latents

Trim a frame range out of an LTX video latent

latent/video
πŸ…›πŸ…£πŸ…§ Set Audio Ref Tokens

Give the model a voice to copy

Lightricks/IC-LoRA
πŸ…›πŸ…£πŸ…§ LTXV Set Audio Video Mask By Time

Time-window inpainting for LTX-2

utility
πŸ…›πŸ…£πŸ…§ LTXV Set Video Latent Noise Masks

Per-frame masks for video inpainting

latent/video
πŸ…›πŸ…£πŸ…§ LTX Sparse Track Editor

Draw motion paths for LTX-2's motion-track control

Lightricks/motion_tracking
πŸ…›πŸ…£πŸ…§ LTXV Spatio Temporal Tiled VAE Decode

Get a long clip out of VRAM jail

latent
πŸ…›πŸ…£πŸ…§ LTXV Stat Norm Latent

Rescale an LTX latent's statistics in one shot

Lightricks/latents
πŸ…›πŸ…£πŸ…§ LTXV Tiled Sampler

High-res LTX video on a smaller card

sampling
πŸ…›πŸ…£πŸ…§ LTXV Tiled VAE Decode

The lighter, spatial-only way to survive the decode step

latent
Modify LTX Model

The model prep node for LTX's editing tricks

ltxtricks
πŸ…›πŸ…£πŸ…§ Multimodal Guider

Separate control over video, audio, and sync

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ Multi Prompt Provider

Batch-encode multiple prompts into one conditioning bundle

prompt
πŸ…›πŸ…£πŸ…§ Set VAE Decoder Noise

The anti-plastic knob for LTX output

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ STG Advanced Presets

The easy button for spatio-temporal guidance

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ STG Guider

The simple sharpen-motion guider for LTX

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ STG Guider Advanced

Per-sigma CFG and STG scheduling

lightricks/LTXV
πŸ…›πŸ…£πŸ…§ STG Guider Node

Spatiotemporal Skip Guidance for sharper, more coherent LTX video

lightricks/LTXV
Readme

ComfyUI-LTXVideo

GitHub Website Model LTXV Trainer Demo Paper Discord

A collection of powerful custom nodes that extend ComfyUI's capabilities for the LTX-2 video generation model.

LTX-2 is built into ComfyUI core (see it here), making it readily accessible to all ComfyUI users. This repository hosts additional nodes and workflows to help you get the most out of LTX-2's advanced features.

To learn more about LTX-2 See the main LTX-2 repository for model details and additional resources.

Prerequisites

Before you begin using an LTX-2 workflow in ComfyUI, make sure you have:

  • ComfyUI installed (Download here](https://www.comfy.org/download)
  • CUDA-compatible GPU with 32GB+ VRAM
  • 100GB+ free disk space for models and cache

Quick Start πŸš€

We recommend using the LTX-2 workflows available in Comfy Manager.

  1. Open ComfyUI
  2. Click the Manager button (or press Ctrl+M)
  3. Select Install Custom Nodes
  4. Search for β€œLTXVideo”
  5. Click Install
  6. Wait for installation to complete
  7. Restart ComfyUI

The nodes will appear in your node menu under the β€œLTXVideo” category. Required models will be downloaded on first use.

Example Workflows

The ComfyUI-LTXVideo installation includes several example workflows. You can see them all at:

ComfyUI/custom_nodes/ComfyUI-LTXVideo/example_workflows/

LTX-2.3 Workflows:

Older Workflows (LTX-2.0):

Union IC-LoRA Model

We introduce a new Union IC-LoRA model that combines depth and edge (canny) control conditions into a single unified LoRA.

Key Features

  • Unified Control: A single LoRA that supports multiple control conditions (depth or edges).
  • Downsampled Latent Processing: The union LoRA operates on a downsampled latent size, which reduces memory usage and significantly speeds up inference while maintaining quality.

How It Works

The union LoRA is trained to understand and respond to both control signals (depth maps and edge maps) within a single model. The model learns to:

  1. Parse multiple conditions: Identify which control signals are present in the input
  2. Process at reduced resolution: Work on downsampled latents to improve efficiency

HDR IC-LoRA

We provide an HDR IC-LoRA that generates linear HDR video encoded in ARRI LogC3, enabling workflows that output high-dynamic-range content suitable for grading and EXR export.

Key Features

  • Linear HDR output: The LoRA produces frames in LogC3-compressed space; the LTXVHDRDecodePostprocess node decodes these back to linear HDR values.
  • SDR preview + raw HDR: The node outputs both a Reinhard-tonemapped SDR preview and the raw linear HDR tensor for downstream use.
  • EXR export: Optionally writes the linear HDR frames as a 16/32-bit EXR image sequence. To enable EXR writing, set OPENCV_IO_ENABLE_OPENEXR=1 in the environment before starting ComfyUI. The exported EXR sequence is best viewed in DJV (or DJV for macOS).

Lipdub IC-LoRA

We provide a Lipdub IC-LoRA that dubs or rephrases speech in video. Given a source video and a text prompt containing the desired dialogue, it generates new lip movements and audio that match the target text while preserving the speaker's identity.

Key Features

  • Multilingual dubbing: Translate speech into another language - the model regenerates lips and audio to match.
  • Same-language rephrasing: Change what the speaker says while keeping the original language.
  • Two-stage pipeline: Stage 1 generates the video and audio at base resolution; Stage 2 upscales while freezing the audio.
  • Speaker identity preservation: Reference audio tokens provide speaker context so the generated voice stays consistent.

Pixel Spatial Upscaler IC-LoRA

We provide Pixel Spatial Upscaler IC-LoRAs that creatively upscale low-resolution video by synthesizing fine detail rather than simply interpolating pixels. Given a low-resolution reference clip, the model re-renders it at 2Γ— or 4Γ— resolution with generative spatial detail β€” making it a creative upsampler, not a pixel-accurate refiner.

Key Features

  • 2Γ— and 4Γ— variants: Choose the 2Γ— upscaler for moderate upscaling or the 4Γ— upscaler for larger resolution jumps.
  • Generative detail synthesis: The model synthesizes texture and structure from the reference rather than faithfully preserving every pixel.
  • Draft-then-upscale workflow: Generate at a low base resolution (e.g. ~280p) to lock in composition and motion, then run the upscaler for the final high-resolution output.
  • Tunable fidelity: LoRA strength, guidance, and step count control how closely the output follows the reference β€” lower values stay closer to the source; higher values allow more creative detail.

Text-to-Audio (T2A)

LTX-2 is a single joint audio/video transformer, but it can generate audio on its own. The LTXVAudioOnlyModel node puts the model into audio-only mode for text-to-audio, with no video output.

Key Features

  • Audio-only sampling: The node sets the model's run_vx, a2v_cross_attn and v2a_cross_attn flags off, so the audio is denoised with no dependence on the video latent and the video stream is skipped. This matches the reference single-stage T2A pipeline's video=None behavior.
  • Minimal dummy video latent: The model splits its input positionally into [video, audio], so the sampler still needs a video latent at index 0. Use the LTXVAudioOnlyEmptyVideoLatent node (a fixed 64x64 single-frame placeholder, no params to tweak) joined with the audio latent via LTXVConcatAVLatent; with LTXVAudioOnlyModel active it is never attended to and adds negligible cost.
  • Audio decode: LTXVAudioVAEDecode extracts the audio directly from the joint latent, then save it with a standard built-in audio node (for example Save Audio (FLAC)).

Required Models

Download the following models:

LTX-2.3 Model Checkpoint - Choose and download one of the models to COMFYUI_ROOT_FOLDER/models/checkpoints folder.

Spatial Upscaler - Required for current two-stage pipeline implementations in this repository. Download to COMFYUI_ROOT_FOLDER/models/latent_upscale_models folder.

Temporal Upscaler - Required for current two-stage pipeline implementations in this repository. Download to COMFYUI_ROOT_FOLDER/models/latent_upscale_models folder.

Distilled LoRA - Required for current two-stage pipeline implementations in this repository (except DistilledPipeline and ICLoraPipeline). Download to COMFYUI_ROOT_FOLDER/models/loras folder.

Gemma Text Encoder Download all files from the repository to COMFYUI_ROOT_FOLDER/models/text_encoders/gemma-3-12b-it-qat-q4_0-unquantized.

LoRAs Choose and download to COMFYUI_ROOT_FOLDER/models/loras folder.

Advanced Techniques

Low VRAM

  • For systems with low VRAM you can use the model loader nodes from low_vram_loaders.py. Those nodes ensure the correct order of execution and perform the model offloading such that generation fits in 32 GB VRAM.
  • Use --reserve-vram ComfyUI parameter: python -m main --reserve-vram 5 (or other number in GB).
  • For complete information about using LTX-2 models, workflows, and nodes in ComfyUI, please visit our Open Source documentation.