Nodes/noEmbryo nodes/H3 AV Latent from Video /noEmbryo
ComfyUI Node

H3 AV Latent from Video /noEmbryo

Encodes a whole video (IMAGE frames + AUDIO) with the MiniMax H3 VAEs into an AV latent that can be saved with 'H3 Motion Context Save Latent' and stitched into later generations.

By noembryo·Created 3 years ago·Updated 4 days ago· 42
H3 AV Latent from Video /noEmbryo
  • video_vae
  • audio_vae
  • images
  • latent
  • audio
  • latent
◄source_fps24.000►
CategorynoEmbryo/MiniMax H3

Inputs (6)

NameTypeDefaultDescription
video_vaeVAEMiniMax H3 video VAE (FP16 or INT8 ConvRot).
audio_vaeVAEMiniMax H3 audio VAE FP32.
source_fpsFLOAT24.0001–240The frame rate of the loaded video. Frames are resampled to H3's native 24 fps by time-based frame picking, so audio stays in sync at any source rate.
imagesoptIMAGEThe whole video as frames (e.g. from VHS Load Video). Leave un-connected when using the latent input.
latentoptLATENTOptional alternative to images. Accepts either a nested AV latent from LTXVConcatAVLatent (its audio stream is used directly) or an already-encoded H3 video LATENT from VAE Encode. A standard [B,C,H,W] latent is wrapped as a one-frame H3 video stream; connect AUDIO separately when it has no audio stream.
audiooptAUDIOThe video's audio (e.g. from VHS Load Video). Leave un-connected for a silent clip. Ignored when latent already contains an audio stream.

Outputs (1)

NameTypeDescription
latentLATENT—