ComfyUI Node
H3 AV Latent from Video /noEmbryo
Encodes a whole video (IMAGE frames + AUDIO) with the MiniMax H3 VAEs into an AV latent that can be saved with 'H3 Motion Context Save Latent' and stitched into later generations.
H3 AV Latent from Video /noEmbryo
- video_vae
- audio_vae
- images
- latent
- audio
- latent
◄source_fps24.000►
CategorynoEmbryo/MiniMax H3
Inputs (6)
| Name | Type | Default | Description |
|---|---|---|---|
| video_vae | VAE | MiniMax H3 video VAE (FP16 or INT8 ConvRot). | |
| audio_vae | VAE | MiniMax H3 audio VAE FP32. | |
| source_fps | FLOAT | 24.0001–240 | The frame rate of the loaded video. Frames are resampled to H3's native 24 fps by time-based frame picking, so audio stays in sync at any source rate. |
| imagesopt | IMAGE | The whole video as frames (e.g. from VHS Load Video). Leave un-connected when using the latent input. | |
| latentopt | LATENT | Optional alternative to images. Accepts either a nested AV latent from LTXVConcatAVLatent (its audio stream is used directly) or an already-encoded H3 video LATENT from VAE Encode. A standard [B,C,H,W] latent is wrapped as a one-frame H3 video stream; connect AUDIO separately when it has no audio stream. | |
| audioopt | AUDIO | The video's audio (e.g. from VHS Load Video). Leave un-connected for a silent clip. Ignored when latent already contains an audio stream. |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| latent | LATENT | — |