Nodes/ComfyUI-H3-Cast/MiniMax H3 Cast to Video
ComfyUI Node

MiniMax H3 Cast to Video

A ComfyUI node in MiniMax H3/cast with 21 inputs and 4 outputs.

By kat3ri·Created 16 days ago·Updated 9 days ago· 2
MiniMax H3 Cast to Video
  • clip
  • vae
  • audio_vae
  • cast_1
  • cast_2
  • cast_3
  • scene_images
  • positive
  • latent
  • final_prompt
  • report
prompt
width1344
height768
length124
max_views_per_member3
auto_introtrue
scene_description
include_voicestrue
ref_image_sizematch
ref_spacing1.0
ref_strength1.00
ref_decay0.00
ref_ramp0.0
temporal_stretch1.0
CategoryMiniMax H3/cast

Inputs (21)

NameTypeDefaultDescription
clipCLIP
vaeVAE
promptSTRINGRefer to characters by NAME -- the <Picture i>/<Audio j> intro lines are written for you (see final_prompt output)
widthINT134432–8192
heightINT76832–8192
lengthINT1245–3600
max_views_per_memberINT31–9Views taken per cast member, in saved order. 9 total image slots are shared by all members + scene images.
auto_introBOOLEANtrueWrite the '<Picture 1>, <Picture 2>: Name -- description' intro lines automatically
audio_vaeoptVAE
cast_1optH3_CAST_MEMBER
cast_2optH3_CAST_MEMBER
cast_3optH3_CAST_MEMBER
scene_imagesoptIMAGEExtra reference images of the location/scene (e.g. room renders); batch = one ref slot per frame
scene_descriptionoptSTRINGWhat's actually in scene_images -- used verbatim in the intro line instead of the generic 'the location where this scene takes place.' placeholder. Concrete detail here (like Plan v2's Image Reference description) is what actually anchors the model to the background; leave blank only if you want the old generic behavior.
include_voicesoptBOOLEANtrue
ref_image_sizeoptCOMBOmatch2 options: match, max
ref_spacingoptFLOAT1.00–50
ref_strengthoptFLOAT1.000–1
ref_decayoptFLOAT0.000–1
ref_rampoptFLOAT0.00–50
temporal_stretchoptFLOAT1.01–100

Outputs (4)

NameTypeDescription
positiveCONDITIONING
latentLATENT
final_promptSTRING
reportSTRING