Nodes/Muse Collective LTX Timeline/Muse Collective LTX Timeline V5 (MSR)
ComfyUI Node

Muse Collective LTX Timeline V5 (MSR)

Subject Reference Video, Composited Into the Timeline

By muse-collective-26·Created 3 months ago·Updated 2 months ago· 9
Muse Collective LTX Timeline V5 (MSR)
  • model
  • clip
  • audio_vae
  • vae
  • spatial_upscaler
  • bg_audio
  • base_model
  • msr_subject_1
  • msr_subject_2
  • msr_subject_3
  • msr_subject_4
  • msr_background
  • last_chunk_frames
  • audio
  • stage1_frames
  • seed_hunt_preview_1
  • seed_hunt_preview_2
  • seed_hunt_preview_3
  • seed_hunt_preview_4
  • seed_hunt_audio_1
  • seed_hunt_audio_2
  • seed_hunt_audio_3
  • seed_hunt_audio_4
◄start_second0.00►
◄end_second10.00►
◄duration_seconds10.00►
◄start_frame0►
◄end_frame240►
◄duration_frames240►
◄timeline_data{}►
◄local_prompts►
◄segment_lengths►
◄global_prompt►
◄guide_strength►
◄epsilon0.0010►
◄frame_rate24.00►
◄display_modeseconds►
◄custom_width960►
◄custom_height544►
◄resize_methodmaintain aspect ratio►
◄divisible_by32►
◄img_compression18►
◄generate_audiotrue►
◄custom_audio_onfalse►
◄lipsynctrue►
◄motion_guide_ontrue►
◄chunk_duration_seconds10.0►
◄auto_chunk_threshold10.0►
◄carry_frames73►
◄carry_strength1.00►
◄crossfade_frames0►
◄ic_lora_nameNone►
◄ic_lora_strength1.00►
◄stage1_steps8►
◄stage2_steps4►
◄stage2_denoise0.42►
◄cfg1.0►
◄seed42►
◄filename_prefixmuse►
◄bg_volume1.00►
◄guide_scale_by0.50►
◄guide_scale_by_s21.00►
◄guide_upscale_methodbicubic►
◄guide_image_attn_strength1.00►
◄guide_cropcenter►
◄guide_auto_snap_ic_gridtrue►
◄guide_use_tiled_encodefalse►
◄guide_tile_size256►
◄guide_tile_overlap64►
◄timeline_ui►
◄seed_huntfalse►
◄seed_hunt_steps6►
◄seed_hunt_scale0.25►
◄seed_hunt_11►
◄seed_hunt_22►
◄seed_hunt_33►
◄seed_hunt_44►
◄use_seed_hunt_1false►
◄use_seed_hunt_2false►
◄use_seed_hunt_3false►
◄use_seed_hunt_4false►
◄msr_frame_count41►
◄msr_strength1.00►

V5 is the MSR version - and MSR stands for Licon's Multi-Subject Reference, one of the LTX 2.3 IC-LoRA tricks that lets a short reference video define who's in the shot and what the camera does. Where V3/V4 lock identity through face or character references, V5 lets you hand the director an actual subject video plus a background video and have it composite them into a guide that steers the whole generation. It's the node you reach for when "reference image" isn't enough and you want motion, angles, and layout carried over.

MSR lives in the LTX 2.3 ecosystem as a family of IC-LoRAs (the LTX-2.3-Licon-MSR-V2.safetensors type files), and the community has been running them at surprisingly low VRAM - there are reports of it working on an 8GB card with the right quants. V5 wires that capability into the Muse director's timeline machinery.

How it works

MSR is driven by a composited reference - your subject clips laid over a background - that gets applied as an IC-LoRA guide. The optional inputs make that explicit:

  • msr_subject_1 … msr_subject_4 - subject reference images/clips (up to four).
  • msr_background - the background reference. Required for MSR to activate. The tooltip is blunt: wiring only subject images does nothing. This is the number-one mistake.
  • msr_frame_count - how long the composited reference runs (17–65 frames, default 41).
  • msr_strength - how strongly the composited reference is applied as an IC-LoRA guide (0–1, default 1).

The critical wiring detail: you must set ic_lora_name to the MSR LoRA. MSR reuses the director's existing IC-LoRA loading path - it doesn't load its own LoRA. No MSR LoRA selected, no MSR effect, even with everything else wired.

Since it's built on V2, you keep the full timeline director: Seed Hunt (seed_hunt, use_seed_hunt_1..4), per-segment prompts, [SPEECH]/[SOUNDS] tags, chunking, two-stage sampling, and the standard outputs plus the four seed-hunt preview/audio pairs.

Installing it

Same pack as ever:

cd ComfyUI/custom_nodes
git clone https://github.com/muse-collective-26/muse-ltx-timeline

Restart, pip install av torchaudio soundfile. For the MSR side you need the MSR LoRA file itself in models/loras/ (the README points at the LTX2.3/ subfolder for the talking-head LoRA; drop the MSR LoRA wherever ComfyUI scans for LoRAs) and the LTX 2.3 stack.

Gotchas

  • MSR won't turn on without msr_background. If your output shows no MSR influence at all, check that socket first, then ic_lora_name.
  • The MSR LoRA is not optional hardware - this is an IC-LoRA feature, and IC-LoRA means model files.
  • msr_frame_count isn't a free knob: short counts read as a flash of reference, long counts cost latency and can over-anchor the shot. 41 (the default) is a fine starting point.
  • WIP module again - loaded in a try/except, so a failed dependency silently drops the node from your list.

V5 is niche but potent: it's the Muse answer to "I want this specific subject, in this specific setting, doing motion consistent with a reference." Set up correctly it's one of the strongest identity-holding options in the pack, and the MSR LoRAs are cheap enough that it's worth an afternoon.

CategoryMuse Collective

Inputs (72)

NameTypeDefaultDescription
modelMODEL—
clipCLIP—
audio_vaeVAE—
vaeVAE—
spatial_upscalerLATENT_UPSCALE_MODEL—
start_secondFLOAT0.000–3600—
end_secondFLOAT10.000–3600—
duration_secondsFLOAT10.000–3600—
start_frameINT00–86400—
end_frameINT2400–86400—
duration_framesINT2401–86400—
timeline_dataSTRING{}—
local_promptsSTRING—
segment_lengthsSTRING—
global_promptSTRING—
guide_strengthSTRING—
epsilonFLOAT0.00100–1—
frame_rateFLOAT24.001–120—
display_modeCOMBOseconds2 options: seconds, frames
custom_widthINT96064–4096—
custom_heightINT54464–4096—
resize_methodCOMBOmaintain aspect ratio4 options: maintain aspect ratio, stretch to fit, crop, pad
divisible_byINT321–256—
img_compressionINT180–51—
generate_audioBOOLEANtrueLTX generates ambient/sfx audio from [SOUNDS] prompts.
custom_audio_onBOOLEANfalseUse audio file(s) from the AUDIO timeline track.
lipsyncBOOLEANtrueSync mouth movements to custom audio. Requires Custom Audio ON and talking head LoRA.
motion_guide_onBOOLEANtrueUse motion guide segments from the timeline.
chunk_duration_secondsFLOAT10.02–120—
auto_chunk_thresholdFLOAT10.00–3600—
carry_framesINT731–240Reference frames from previous chunk locked at chunk start. 73 ≈ 3s at 24fps.
carry_strengthFLOAT1.000–1—
crossfade_framesINT00–120—
ic_lora_nameCOMBONone1 options: None
ic_lora_strengthFLOAT1.00-10–10—
stage1_stepsINT81–50—
stage2_stepsINT41–50—
stage2_denoiseFLOAT0.420–1—
cfgFLOAT1.00–20—
seedINT420–18446744073709550000—
filename_prefixSTRINGmuse—
bg_volumeFLOAT1.000–2—
guide_scale_byFLOAT0.500.01–8—
guide_scale_by_s2FLOAT1.000.01–8—
guide_upscale_methodCOMBObicubic5 options: bicubic, bilinear, nearest-exact, area, bislerp
guide_image_attn_strengthFLOAT1.000–1—
guide_cropCOMBOcenter2 options: center, disabled
guide_auto_snap_ic_gridBOOLEANtrue—
guide_use_tiled_encodeBOOLEANfalse—
guide_tile_sizeINT25664–512—
guide_tile_overlapINT6416–256—
timeline_uiSTRING—
seed_huntBOOLEANfalseON + no candidate chosen: run a 4-seed Stage-1-resolution preview instead of the full pipeline. ON + one use_seed_hunt_N chosen: commit to that candidate — Stage 2 refines its actual cached latent instead of regenerating Stage 1 from scratch.
seed_hunt_stepsINT61–50—
seed_hunt_scaleFLOAT0.250.05–1Unused as of 1.0.4 — Seed Hunt now scouts at Stage 1's real resolution automatically (so the picked candidate's actual latent can carry forward into Stage 2). Kept as a widget only so older saved workflows still load correctly.
seed_hunt_1INT10–18446744073709550000Unused as of 1.0.4 — scouting now draws a fresh random seed for each candidate every run instead of reusing these fixed values (the actual latent carries forward on commit, so the seed number no longer needs to be fixed or reproducible).
seed_hunt_2INT20–18446744073709550000—
seed_hunt_3INT30–18446744073709550000—
seed_hunt_4INT40–18446744073709550000—
use_seed_hunt_1BOOLEANfalse—
use_seed_hunt_2BOOLEANfalse—
use_seed_hunt_3BOOLEANfalse—
use_seed_hunt_4BOOLEANfalse—
bg_audiooptAUDIO—
base_modeloptMODELBase model without talking-head LoRA. Connect the UNETLoader output directly here so the ambient audio pass generates sounds without speech.
msr_subject_1optIMAGEMSR subject reference 1 (optional). Requires msr_background to be wired for MSR to activate.
msr_subject_2optIMAGEMSR subject reference 2 (optional).
msr_subject_3optIMAGEMSR subject reference 3 (optional).
msr_subject_4optIMAGEMSR subject reference 4 (optional).
msr_backgroundoptIMAGEMSR background reference. Required to activate MSR — wiring only subject images does nothing.
msr_frame_countoptCOMBO417 options: 17, 25, 33, 41, 49, 57, +1
msr_strengthoptFLOAT1.000–1How strongly the composited MSR reference is applied as an IC-LoRA guide. Requires ic_lora_name set to the MSR LoRA (e.g. LTX-2.3-Licon-MSR-V2.safetensors) — MSR reuses the same IC-LoRA loading Director already has, it does not load its own LoRA.

Outputs (11)

NameTypeDescription
last_chunk_framesIMAGE—
audioAUDIO—
stage1_framesIMAGE—
seed_hunt_preview_1IMAGE—
seed_hunt_preview_2IMAGE—
seed_hunt_preview_3IMAGE—
seed_hunt_preview_4IMAGE—
seed_hunt_audio_1AUDIO—
seed_hunt_audio_2AUDIO—
seed_hunt_audio_3AUDIO—
seed_hunt_audio_4AUDIO—