Nodes/ComfyUI-Model-Bending/Read Attention Maps (Experimental)
ComfyUI Node

Read Attention Maps (Experimental)

Turns recorded attention maps into video frames (IMAGE batch): bright = the token attends there. Experimental: validated only on tiny random-weight WAN models in the test suite, not yet on real video model weights.

By abuzreq·Created 2 years ago·Updated 2 days ago· 23
Read Attention Maps (Experimental)
  • attention_maps
  • latent
  • frames
  • report
◄tokensum►
◄blockmean►
◄stepall►
◄normalizeper_video►
◄colormapinferno►
◄match_video_framestrue►
Categorymodel_bending/video (experimental)

Inputs (8)

NameTypeDefaultDescription
attention_mapsATTENTION_MAPS—
latentLATENTConnect the output of the sampler that used the capture model, so this node runs after sampling
tokenSTRINGsum'sum' of the recorded tokens, or a token index
blockSTRINGmean'mean' over recorded blocks, or a block index
stepSTRINGall'all' = every recorded step side by side, 'mean', or a step index
normalizeCOMBOper_video2 options: per_video, per_frame
colormapCOMBOinferno2 options: inferno, gray
match_video_framesBOOLEANtrueRepeat each latent frame 4x (after the first) so the maps line up with WAN's decoded frames

Outputs (2)

NameTypeDescription
framesIMAGE—
reportSTRING—