BFS Shot H3 Duet (render one shot)
Renders one shot with MiniMax H3 (duet, aligned guide, or both) and returns its frames for BFS Shot Join. Runs once per shot of the Planner's list. The prompt comes from the Planner (per shot or global), or from the task when there is none. Duet modes: How to refer to the side panel in the prompt: the panel has NO tag (it is not <Picture n> or <Video n>; the text encoder never sees it). Name it by its place: 'the LEFT half is the kept footage' (write {layout} to insert that sentence for the current size and side). Describe only the generated part: never describe the panel's performer, clothes or room, even to contrast them; what you leave undescribed is copied from the panel. Restate the new identity in every shot ('her face from <Picture 1>' plus two or three face, hair or outfit words). Give exact times for cuts ('[Shot 2] At 00:03.708, both halves cut together to ...'). Credits: the initial idea for this H3 node came from TSC's latent-pin duet (the source pinned beside the video with a noise mask). BFS had already used the same principle on LTX (a green side panel holding the reference) and implemented the virtual sidecar approach (reference tokens placed beside the frame in RoPE). New here: the shifted RoPE layout (the video keeps its own grid and the panel sits past its edge, with an optional gap), dynamic references, task prompts and the shot-loop node.
- shot
- model
- clip
- vae
- audio_vae
- images
- audio
- canvas
- layout_text
- prompt
Inputs (22)
| Name | Type | Default | Description |
|---|---|---|---|
| shot | BFS_SHOT | — | |
| model | MODEL | — | |
| clip | CLIP | — | |
| vae | VAE | — | |
| mode | COMBO | duet (pin the shot, no LoRA needed) | duet: the shot's own clip is pinned beside the video and copied in sync. guide: the shot sits on the generated frames as a latent guide (use with a body-swap LoRA). duet + guide: both. |
| task | COMBO | character swap | Writes the prompt when the shot has none (no per-shot or global prompt in the Planner). |
| instruction | STRING | What changes, in a few seen words (see BFS H3 Duet). | |
| use_ref_2 | BOOLEAN | true | — |
| position | COMBO | left | Side of the video the strip is added to (TSC's duet pins the source clip on the left). |
| size | FLOAT | 1.000.1–1.5 | Strip size as a fraction of the video's height (top/bottom) or width (left/right), snapped to 32 px. |
| fit | COMBO | contain | contain keeps the whole clip (smaller, with grey around) so nothing is cropped |
| gap | INT | 00–8 | Gray separator between panel and video, in 32 px patches (held like the panel). |
| panel_noise | FLOAT | 0.000–1 | 0 pins the panel exactly. 0.05-0.15 lets the model loosen it a little when the result copies too much of it (TSC's SOURCE NOISE). |
| ref_image_size | COMBO | match | 2 options: match, max |
| steps | INT | 201–200 | — |
| sampler_name | COMBO | euler | 44 options: euler, euler_cfg_pp, euler_ancestral, euler_ancestral_cfg_pp, heun, heunpp2, +38 |
| scheduler | COMBO | beta | 9 options: simple, sgm_uniform, karras, exponential, ddim_uniform, beta, +3 |
| seed | INT | 420–18446744073709550000 | — |
| decode_canvas | BOOLEAN | false | — |
| audio_vaeopt | VAE | — | |
| rope_modeopt | COMBO | canvas | 2 options: canvas, shifted |
| rope_gapopt | FLOAT | 00–256 | — |
Outputs (5)
| Name | Type | Description |
|---|---|---|
| images | IMAGE | — |
| audio | AUDIO | — |
| canvas | IMAGE | — |
| layout_text | STRING | — |
| prompt | STRING | — |