Extensions/ComfyUI
ComfyUI Extension Runs on cloud

ComfyUI

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

By Comfy-Org·Created 4 years ago·Updated about 11 hours ago· 128,195
comfyanonymous/ComfyUI
Nodes848
On cloudRunnable
Categoryimage, model/sampling/noise
Stars128,195
Updatedabout 11 hours ago

Nodes (848)

Add Layer

Assemble a scene without flattening it

image
AddNoise

Add noise to a latent by hand — the manual version of what samplers do silently

model/sampling/noise
Add Text Prefix (DEPRECATED)

A deprecated node you can safely skip

text
Add Text Suffix (DEPRECATED)

The deprecated mirror of Add Text Prefix

text
Adjust Brightness

One multiply, and the clamp that eats your highlights

image/adjustments
Adjust Contrast

The one-knob contrast slider hiding inside ComfyUI's training pipeline

image/adjustments
AlignYourStepsScheduler

NVIDIA's 10-step presets, still quietly useful

model/sampling/schedulers
Apply Anima LLLite

The closest thing Anima has to a ControlNet

model_patches/anima
Adaptive Projected Guidance

Adaptive Projected Guidance, the guidance hack from the audio world

model/sampling/custom
ARVideoI2V

Seed an autoregressive video model with a start frame

model/conditioning/autoregressive
Adjust Audio Volume

The difference between 'mixed' and 'buried'

audio
Concatenate Audio

Splicing audio the way you already splice images

audio
AudioEncoderEncode

The generic bridge between audio and its latent form

model/conditioning
Load Audio Encoder

How ComfyUI hears your video's soundtrack

model/loaders
Audio Equalizer (3-Band)

A real EQ in your graph, marked experimental for good reason

audio
Merge Audio

The mixer that treats audio like a canvas

audio
Basic Guider

No negative, no CFG, no drama

model/sampling/guiders
BasicScheduler

The noise schedule, as a SIGMAS object

model/sampling/schedulers
Batch Images

The node that replaced the deprecated one

image/batch
Batch Latents

Gather many latents into one stack

model/latent/batch
Batch Masks

Stack a pile of masks into one tensor and get a free batch

image/mask
Beeble SwitchX Image Edit

Change the scene, keep the subject's pixels

partner/image/Beeble
Beeble SwitchX Video Edit

Swap the world around a subject, keep the subject's pixels

partner/video/Beeble
Bernini Conditioning

Bernini's whole bag of tricks in one node

model/conditioning/bernini
BetaSamplingScheduler

The gentle schedule that flow-matching models actually like

model/sampling/schedulers
Bria Eraser

Remove anything under a mask and let the cloud fill the hole — Bria Eraser

partner/image/Bria
Bria Expand Image
partner/image/Bria
Bria Generative Fill

Add objects to a masked region with a prompt — Bria GenFill, the 'generate, don't erase' sibling of Eraser

partner/image/Bria
Bria FIBO Image Edit

The masked-edit node that hands you back its own prompt

partner/image/Bria
Bria Increase Resolution
partner/image/Bria
Bria Remove Image Background

Background removal as a paid API call — is that ever worth it?

partner/image/Bria
Bria Remove Video Background

The hosted cutter for moving subjects

partner/video/Bria
Bria Remove Video Background (Transparent)

Cut a subject out of video and keep the alpha — Bria transparent background

partner/video/Bria
Bria Video Green Screen

Bake in a real chroma-key background, so the keying happens later

partner/video/Bria
Bria Video Replace Background

Swap a video's background for an image or another video (Bria)

partner/video/Bria
Build JSON Prompt (Ideogram)

Layout control for the text-in-image king

text
ByteDance Seedance 2.5 First-Last-Frame to Video

Pin the beginning and end, let Seedance 2.0 fill in the middle

partner/video/ByteDance
ByteDance Seedance 2.5 Reference to Video

The multi-input reference node that finally makes 'that same person' work across shots

partner/video/ByteDance
ByteDance Seedance 2.5 Text to Video

Seedance 2.0 text to video, with dialogue steering and a Fast/Mini cost ladder

partner/video/ByteDance
ByteDance Create Image Asset

Register your face so Seedance can reuse it — the asset node

partner/image/ByteDance
ByteDance Create Video Asset

Register a personal Seedance video asset (once) so every future gen matches

partner/video/ByteDance
ByteDance First-Last-Frame to Video

ByteDance's first-to-last-frame video

partner/video/ByteDance
ByteDance Image

ByteDance's Seedream 3.0 in ComfyUI — the retired version

partner/image/ByteDance
ByteDance Reference Images to Video

Seedance 1.0, up to four images, and one naming trap

partner/video/ByteDance
ByteDance Image to Video

Animate your first frame with Seedance 1.x, then hold the camera still if you like

partner/video/ByteDance
ByteDance Seed Audio 1.0

Voice, music, SFX, and dialogue

partner/audio/ByteDance
ByteDance Seed

A multimodal LLM that will happily watch your video clips

partner/text/ByteDance
ByteDance Seedream 5.0 Pro Layer Separation

One image in, a PSD-style layer stack out

partner/image/ByteDance
ByteDance Seedream 4.5 & 5.0

The closed flagship image API, now one node away

partner/image/ByteDance
ByteDance Seedream 4.5 & 5.0

Seedream up to 4K — ByteDance's API-only flagship, wired into your graph

partner/image/ByteDance
ByteDance Text to Video

Prompt, resolution, duration — done

partner/video/ByteDance
Detect Edges (Canny)

The edge map that started ControlNet

image/filters
Convert Text Case

Four ways to yell or whisper at your prompt

text
Crop Image (Center)

The boring, reliable way to square up a batch

image/transform
CFG Guider

Your prompt, turned into guidance, as an object

model/sampling/guiders
CFGNorm

Keep high CFG from deep-frying your colors

advanced/guidance
CFG Override

Different CFG for different parts of the run

model/sampling/guiders
CFGZeroStar

Crank CFG hard without the burn

advanced/guidance
Load Checkpoint With Config (DEPRECATED)

The old checkpoint loader that needed a config file — and why it died

model/loaders
Load Checkpoint

The beginner's one-stop shop that modern workflows outgrew

model/loaders
Save Checkpoint

The node that turns your merge into a file (and hides it from you)

model/merging
ChromaRadianceOptions

The advanced-options node you'll only touch for one thing

model/patch/chroma radiance
Anthropic Claude

Reasoning budgets and a hard 20-image cap

partner/text/Anthropic
CLIPAttentionMultiply

Scaling CLIP attention before it reads your prompt

experimental/attention_experiments
Load CLIP

The dropdown that decides whether your prompt is even readable

model/loaders
CLIPMergeAdd

Adding text encoders together, the bluntest of the CLIP merges

model/merging
CLIPMergeSimple

Blending text encoders, and the encoder-swap trick people actually use

model/merging
CLIPMergeSubtract

The text-encoder difference node for de-training a bad CLIP

model/merging
CLIPSave

The node that saves a text encoder (and splits it into pieces on the way)

model/merging
CLIP Set Last Layer

The CLIP-skip dial, and the three models where it actually changes anything

model/conditioning
CLIP Text Encode (Prompt)

The node your whole workflow starts at

model/conditioning
CLIP Text Encode (Controlnet)

Give your ControlNet its own separate prompt

model/conditioning
CLIPTextEncodeFlux

Two prompt boxes for two encoders, and the guidance dial everyone fights over

model/conditioning/flux
CLIP Text Encode (HiDream)

HiDream-I1's text path

model/conditioning/hidream
CLIP Text Encode (Hunyuan Image)

Prompt Hunyuan Image in English or Chinese

model/conditioning/hunyuan image
CLIP Text Encode (Kandinsky 5)

Kandinsky 5's CLIP + Qwen prompt node

model/conditioning/kandinsky
CLIP Text Encode (Lumina 2)

The LLM-style prompt box with a system-prompt dropdown

model/conditioning/lumina
CLIP Text Encode (PixArt Alpha)

PixArt Alpha's text encoder with a resolution hidden inside

model/conditioning/pixart
CLIP Text Encode (SD3)

Prompting SD3 and SD3.5

model/conditioning/stable diffusion
CLIP Text Encode (SDXL)

The two-prompt encoder SDXL actually wants

model/conditioning/stable diffusion
CLIP Text Encode (SDXL Refiner)

The aesthetic-score prompt box for the model everyone decided to skip

model/conditioning/stable diffusion
CLIP Vision Encode

The node that lets your image talk

model/conditioning
Load CLIP Vision

The encoder that lets a model see your image

model/loaders
Color Picker

Pick a color, get a number — for the nodes that want RGB as an int

utilities
Transfer Color

Borrow any photo's look with one wire

image/filters
Combine Hooks [2]

The two-input switchboard for your hook groups

advanced/hooks/combine
Combine Hooks [4]

A four-way merge for real multi-region workflows

advanced/hooks/combine
Combine Hooks [8]

The biggest switchboard in the hook family

advanced/hooks/combine
And

One gate for every condition you're testing

utilities/logic
Math Expression

A calculator that lives inside your graph — type the formula, get three outputs

utilities
Not

Flip any value into its opposite

utilities/logic
Convert Number

One node to cast anything to a number — int, float, string, bool

utilities
Or

True if anything in the pile is true

utilities/logic
If/Else Switch

The conditional node ComfyUI finally shipped

utilities/logic
Conditioning (Average)

Blend two prompts into one conditioning

model/conditioning/transform
Conditioning (Combine)

Run two prompts at once, and stack your regional and ControlNet branches

model/conditioning/transform
Conditioning (Concat)

How to bolt a second prompt onto the end of the first

model/conditioning/transform
Conditioning (Multiply)

A volume knob for your prompt

model/conditioning/transform
Conditioning (Set Area)

Regional prompting for two characters that keep bleeding into each other

model/conditioning/transform
Conditioning (Set Area with Percentage)

Regional prompts that survive a resolution change

model/conditioning/transform
Conditioning (Set Area with Percentage for Video)

Regional prompting that also reaches across time

model/conditioning/transform
Conditioning (Set Area Strength)

Change the gain, not the box

model/conditioning/transform
Cond Set Default Combine

Give the empty space in your regional image a prompt

advanced/hooks/cond single
Conditioning (Set Mask)

Make one prompt own one part of the frame

model/conditioning/transform
Cond Set Props

The gateway node that actually hands your hooks to the sampler

advanced/hooks/cond single
Cond Set Props Combine

Apply a hook, mask, and timestep window — then merge, in one node

advanced/hooks/cond single
ConditioningSetTimestepRange

Make a prompt only steer part of the denoise

model/conditioning/transform
ConditioningStableAudio

Telling a music model when in the song you are

model/conditioning/stable audio
Timesteps Range

The when dial for hook-based conditioning

advanced/hooks
Conditioning Zero Out

The explicit empty prompt

model/conditioning/transform
Context Windows (Manual)

Sample long video without your VRAM collapsing

model/patch
Apply ControlNet (DEPRECATED)

The old ControlNet node — use ControlNetApplyAdvanced instead

model/conditioning/controlnet
Apply ControlNet

The node that pins structure to your image, and the start/end dials that make it behave

model/conditioning/controlnet
Apply Controlnet with VAE

The deprecated SD3 applier you'll still find in old workflows

model/conditioning/controlnet
Apply ControlNet Inpainting (AliMama)

The inpainting ControlNet that hides a mask in a channel

model/conditioning/controlnet
Load ControlNet Model

Where structure enters your generation

model/loaders
Convert Array to String

Serialize a list into JSON text

text
Convert Dictionary to String

Turn a JSON object back into text

text
CosmosImageToVideoLatent

Start frame, end frame, and the NVIDIA world-model video it belongs to

model/conditioning/cosmos
CosmosPredict2ImageToVideoLatent

Cosmos-Predict2 image-to-video latent prep

model/conditioning/cosmos
Create Bounding Boxes

Draw boxes on a canvas, get Ideogram-style prompt elements out

utilities
Create Camera Info

The turntable camera that makes splats spin

3d
Create Hook Keyframe

Place one strength stop on your hook's timeline

advanced/hooks/scheduling
Create Hook Keyframes From Floats

Your exact strength curve, fed as raw numbers

advanced/hooks/scheduling
Create Hook Keyframes Interp.

The one node for smooth strength ramps on your LoRA

advanced/hooks/scheduling
Create Hook LoRA

The node that turns a LoRA file into a schedulable, maskable patch

advanced/hooks/create
Create Hook LoRA (MO)

The LoRA hook without the CLIP half

advanced/hooks/create
Create Hook Model as LoRA

Run a whole checkpoint as a maskable, schedulable weight patch

advanced/hooks/create
Create Hook Model as LoRA (MO)

The whole-model patch with the text encoder left out

advanced/hooks/create
Create List

Gather many things into one list — the iterator's supply line

utilities
Create Video

Pack your frames (and audio) back into a video, no extension pack needed

video
Crop By Bounding Boxes

Crop everything a detector found, all at once

image/transform
Crop Mask

Cut a region out of a mask, the honest way

image/mask
Curve Editor

Draw a ramp, not a number — the curve node that comes with a histogram view

utilities
Custom Combo

A dropdown you write yourself — and an index for your switches

utilities
Convert DA3 Geometry to Mesh

From Depth Anything 3 depth to a textured mesh in one node

image/geometry estimation
Run Depth Anything 3

The node that does depth, multi-view consistency, and camera poses

image/geometry estimation
Render Depth Anything 3

Depth, confidence, and sky maps from Depth Anything 3, in colors you can read

image/geometry estimation
Load ControlNet Model (diff)

The variant that wants your base model too

model/loaders
Differential Diffusion

The fix for two-tone seams in masked inpainting

experimental
Load Diffusers Model (DEPRECATED)

The loader for Hugging Face's diffusers format — deprecated, but not useless

model/loaders
DisableNoise

A noise source that adds nothing

model/sampling/noise
Draw BBoxes

The one node that shows you what your detector actually found

image/detection
Dual CFG Guider

Two prompts, two CFG dials, one pass

model/sampling/guiders
Load CLIP (Dual)

The two-encoder loader behind SDXL and Flux

model/loaders
Dual Model CFG Guider

One model for the prompt, another for the blank

model/sampling/guiders
EasyCache

EasyCache — ComfyUI's built-in step-skipper that can nearly halve sampling time

advanced/debug
ElevenLabs Voice Isolation

The voice-isolation node

partner/audio/ElevenLabs
ElevenLabs Instant Voice Clone

Clone a voice from a few samples — your voice, their GPU

partner/audio/ElevenLabs
ElevenLabs Speech to Speech

Take any voice, speak it with any other voice

partner/audio/ElevenLabs
ElevenLabs Speech to Text

Transcribe audio, speakers and sound effects included

partner/audio/ElevenLabs
ElevenLabs Text to Dialogue

Write a script, cast voices, get a finished conversation

partner/audio/ElevenLabs
ElevenLabs Text to Sound Effects

Type 'creaky door' and get a creaky door

partner/audio/ElevenLabs
ElevenLabs Text to Speech

ElevenLabs-grade voices, wired straight into your graph

partner/audio/ElevenLabs
ElevenLabs Voice Selector

Pick a voice from a dropdown — and it's actually free

partner/audio/ElevenLabs
Empty Ace Step 1.5 Latent Audio

Ace Step 1.5's audio canvas — wider channels, faster clock

model/latent/ace
Empty Ace Step 1.0 Latent Audio

The Ace Step 1.0 music canvas — a latent shaped like a song

model/latent/ace
EmptyARVideoLatent

Frames one at a time

model/latent/autoregressive
Empty Audio

The silence node that's more useful than it sounds

audio
EmptyChromaRadianceLatentImage

The pixel-space Chroma blank

model/latent/chroma radiance
EmptyCosmosLatentVideo

The blank reel for NVIDIA's Cosmos world models

model/latent/cosmos
Empty Flux 2 Latent

The blank canvas for every FLUX.2 workflow

model/latent/flux
Empty HiDream-O1 Latent Image

The pixel-space blank that isn't a VAE thing

model/latent/hidream
EmptyHunyuanImageLatent

The blank for Tencent's 80B image MoE

model/latent/hunyuan image
Empty HunyuanVideo 1.0 Latent

The starting line for Tencent's video model

model/latent/hunyuan video
Empty HunyuanVideo 1.5 Latent

HunyuanVideo 1.5's canvas — 32 channels at 16x downscale

model/latent/hunyuan video
Empty Image

A solid-color canvas in pixel space

image
Empty Latent Audio

The generic audio canvas for video models that make sound

model/latent
EmptyLatentHunyuan3Dv2

A flat box holding a volume

model/latent/hunyuan 3d
Empty Latent Image

The blank canvas every txt2img graph starts with

model/latent
EmptyLTXVLatentVideo

The fast-canvas node that starts every LTX clip

model/latent/ltxv
Empty MiniMax H3 AV Latent

The blank canvas MiniMax H3 fills with picture and sound at once

model/latent/minimax
Empty MiniMax Music3 Latent Audio
model/latent/minimax music
EmptyMochiLatentVideo

The latent whose length math is a trap

model/latent/mochi
Empty Qwen Image Layered Latent

The blank canvas for Qwen's RGBA layers

model/latent/qwen
EmptySD3LatentImage

The right blank canvas for SD3 and Flux-family models

model/latent/stable diffusion
Epsilon Scaling

A one-percent nudge that fixes exposure bias

model/patch/unet
ExponentialScheduler

The old SD workhorse that flow-matching models will punish you for

model/sampling/schedulers
ExtendIntermediateSigmas

The node that outsmarted Ideogram's safety filter

model/sampling/sigmas
Feather Mask

The soft edge that keeps inpaints from looking glued on

image/mask
Get Splat

Turn a .ply or .spz file on disk into a splat you can actually use

3d/splat
FlipSigmas

Run a noise schedule backwards, because sometimes the answer is reverse

model/sampling/sigmas
Flux.2 Image

Flux.2 [pro] and [max], one node, no local GPU

partner/image/BFL
Flux.2 [max] Image

Flux.2 [max], the premium API tier with explicit control

partner/image/BFL
Flux.2 [pro] Image

Flux 2 without the 24GB — the hosted [pro] model, deprecated but working

partner/image/BFL
Flux2Scheduler

The resolution-aware schedule that Flux 2 expects you to use

model/sampling/schedulers
Flux 3 Image to Video

Turn up to ten images into one FLUX 3 clip

partner/video/BFL
Flux 3 Text to Video

Text to video with sound, minus the GPU

partner/video/BFL
Flux 3 Video Continuation

FLUX 3 Video Continuation

partner/video/BFL
FluxDisableGuidance

Turn off Flux's guidance embed entirely

model/conditioning/flux
Flux Erase Image

Flux Erase

partner/image/BFL
FluxGuidance

One slider that changes how Flux obeys you

model/conditioning/flux
FluxKontextImageScale

Resize your Kontext reference images to a size the model likes

model/conditioning/flux
Flux.1 Kontext [max] Image

Flux.1 Kontext [max]

partner/image/BFL
Edit Model Reference Method

The obscure switch that fixes broken multi-reference edits

model/conditioning/flux
Flux.1 Kontext [pro] Image

The API tier of the model that made instruction-editing normal

partner/image/BFL
Flux KV Cache

Stop re-encoding your reference image every step

experimental
Flux.1 Expand Image

Flux.1 Expand

partner/image/BFL
Flux.1 Fill Image

Flux.1 Fill via API

partner/image/BFL
Flux 1.1 [pro] Ultra Image

The API's premium tier, minus the prompt fuss

partner/image/BFL
Flux Virtual Try-On

Virtual try-on with two image inputs and zero VRAM

partner/image/BFL
Run Frame Interpolation Model

Use RIFE (or FILM) to smooth out that choppy AI video

video
Load Frame Interpolation Model

The slow-motion and frame-doubling engine

model/loaders
FreeU

The 2023 sharpening trick that most modern models don't want anymore

model/patch/unet
FreeU_V2

The SDXL-friendly second version that fixed FreeU's halo problem

model/patch/unet
FreSca

Tune your guidance in frequency space (and stop oversaturating)

experimental
Nano Banana Pro (Google Gemini Image)

Google's 4K image model, minus the ImageFX tab

partner/image/Gemini
Nano Banana (Google Gemini Image)

The cheap Gemini image node that made a banana famous

partner/image/Gemini
Gemini Input Files

Feeding Gemini your documents — and why its token meter is worth watching

partner/text/Gemini
Nano Banana 2

Nano Banana 2 via ComfyUI — the original node (now deprecated)

partner/image/Gemini
Nano Banana 2

Nano Banana 2, the V2 that finally got a good interface

partner/image/Gemini
Google Gemini

Still works, but it's the deprecated one now

partner/text/Gemini
Google Gemini

Thinking levels and a real token budget

partner/text/Gemini
Google Gemini Omni (Video)

Video plus audio, written like a shopping list

partner/video/Gemini
Generate Video Tracks

Draw a motion path and get Wan-Move tracks

model/conditioning/wan/move
Get IC-LoRA Parameters

The metadata reader that stops your LTXV guides from misaligning

model/conditioning/ltxv
Get Image Size

Get Image Size

image
Get Splat Count

A tiny utility that tells you how many gaussians you're dealing with

3d/splat
Get Video Components

The demux that turns a video file into frames, audio, and the numbers you need

video
GITSScheduler

The 2024 hype scheduler that mostly faded — and what it's still for

model/sampling/schedulers
Load GLIGEN Model

The forgotten way to put things exactly where you said

model/loaders
Apply GLIGEN Text Box

The original 'put this prompt in a box' node, still in core

model/conditioning/gligen
GLSL Shader

A mini Shadertoy living inside your ComfyUI graph

image/shader
Grok Image Edit

Grok's image editor in ComfyUI — the original, and it's retired

partner/image/Grok
Grok Image Edit

Grok Image Edit, the version you should actually be using

partner/image/Grok
Grok Image

XAI's generator, called straight from your ComfyUI graph

partner/image/Grok
Grok Video Edit

Tell Grok to change your video, with a sentence instead of a timeline

partner/video/Grok
Grok Video Extend

Extend a video with Grok Video Extend

partner/video/Grok
Grok Video

Grok video from a prompt or an image, without leaving ComfyUI

partner/video/Grok
Grok Reference-to-Video

Grok video steered by reference images and preset voices

partner/video/Grok
Grow Mask

The padding your inpaints never knew they needed

image/mask
HappyHorse Image to Video

Animate a first frame with HappyHorse, and let the image set the aspect ratio

partner/video/Wan
HappyHorse Reference to Video

Keep a character consistent with HappyHorse reference-to-video

partner/video/Wan
HappyHorse Text to Video

HappyHorse text to video — Wan in the cloud, English or Chinese prompts

partner/video/Wan
HappyHorse Video Edit

An Alibaba edit model that no one can download

partner/video/Wan
HeyGen Avatar Video

A talking presenter from an avatar, no camera needed

partner/video/HeyGen
HeyGen Create Avatar

Make your own HeyGen avatar (and keep the ID)

partner/video/HeyGen
HeyGen Talking Photo

Make any portrait talk

partner/video/HeyGen
HeyGen Text to Speech

A narrator's worth of voices without leaving ComfyUI

partner/audio/HeyGen
HeyGen Video Translate

Dubbed video that keeps the speaker's voice

partner/video/HeyGen
HiDream-O1 Patch Seam Smoothing

The fix for grid lines on a 2048px pixel-space render

model/patch/hidream
HiDream-O1 Reference Images

Attach 1 to 10 reference images to HiDream-O1

model/conditioning/hidream
HitPaw General Image Enhance

A generative upscaler that lives in the cloud

partner/image/HitPaw
HitPaw Video Enhance

HitPaw Video Enhance — upscale a clip without renting a GPU

partner/video/HitPaw
Hunyuan3Dv2Conditioning

Turn a CLIP vision pass into Hunyuan3D-2 conditioning

model/conditioning/hunyuan 3d
Hunyuan3Dv2ConditioningMultiView

Sell the 3D model four angles so it doesn't invent the back

model/conditioning/hunyuan 3d
HunyuanImageToVideo

The I2V setup node for the model the community left behind

model/conditioning/hunyuan video
Hunyuan Latent Refiner

Bridge Hunyuan Video's base pass to its refiner

model/conditioning/hunyuan video
HunyuanVideo15ImageToVideo

Hunyuan Video 1.5's image-to-video setup node

model/conditioning/hunyuan video
Hunyuan Video 15 Latent Upscale With Model

1.5's built-in super-resolution pass

model/latent/hunyhuan video
Hunyuan Video 1.5 Super Resolution

Hunyuan Video 1.5's hidden upscale stage, demystified

model/conditioning/hunyuan video
Load Hypernetwork

The tech LoRA killed, still in your node list

model/loaders
HyperTile

Sample big SDXL latents without OOMing your card

model/patch/unet
Ideogram 4 Scheduler

The resolution-aware schedule Ideogram shipped with

model/sampling/schedulers
Ideogram P-Image

The fast Ideogram that's actually good at text

partner/image/Ideogram
Ideogram V3

Text-to-image, mask editing, and character reference in one cloud node

partner/image/Ideogram
Ideogram V4

The current Ideogram node, and why text-in-image still matters

partner/image/Ideogram
Add Noise to Image

Film grain as a single node

image/filters
Batch Images (DEPRECATED)

The two-image glue you'll stop using

image/batch
Blend Images

The fade knob for two generations

image/filters
Blur Image

The smoothing node your workflow is quietly missing

image/filters
Convert Image Color to Mask

Chroma key without the greenscreen

image/mask
Compare Images

The slider that settles every A/B test

image
Image Composite Masked

Paste one image onto another, mask and all — ComfyUI's compositing workhorse

image/compositing
Create Layered Image

The closest ComfyUI gets to a drag-and-drop Photoshop canvas

image
Crop Image (DEPRECATED)

Crop Image (ImageCrop) — the deprecated Essentials node that moved into core

image/transform
Crop Image

Crop by a bounding box instead of four numbers

image/transform
Deduplicate Images

A handy filter with a very blunt knife

image/batch
Flip Image

Mirror an image horizontally or vertically

image/transform
Get Image from Batch

Get Image from Batch

image/batch
Make Image Grid

Your contact sheet, built in

image/batch
Image Histogram

See an image's tonality as data — 256 bins of light, per channel and overall

utilities
Invert Image Colors

A one-line color flip, plus the alpha gotcha nobody mentions

image/color
Merge List of Tiles to Image

Stitch tiles back without the seams

image/batch
Load Checkpoint Image Only (img2vid model)

The img2vid loader that skips the text encoder

model/loaders
ImageOnlyCheckpointSave

Saving the img2vid checkpoints that condition on images, not text

model/merging
Pad Image for Outpainting

ImagePadForOutpaint pads the canvas and hands you the mask

image/transform
Quantize Image

Posterize anything into a retro palette

image/filters
Image RGB to YUV

Split luminance from color — grayscale, color grading, and the swap trick

image/color
Rotate Image

The 90-degree node with no surprises

image/transform
Upscale Image

The exact-size resize that's still here for a reason

image/upscaling
Upscale Image By

The 'just make it bigger' node

image/upscaling
Scale Image to Max Dimension

The fit-in-a-box resize

image/upscaling
Scale Image to Total Pixels

The resize node hiding in every SeedVR2 recipe

image/upscaling
Sharpen Image

The final-touch node that can't create detail

image/filters
Stitch Images

Put two images side by side, spacer bar optional

image/transform
Convert Image to Mask

Pull one channel out of an image and call it a mask

image/mask
Upscale Image (using Model)

Run an ESRGAN upscaler on your image

image/upscaling
Image YUV to RGB

The other half of ComfyUI's luminance-and-color trick

image/color
InpaintModelConditioning

The node that makes inpainting respect the mask — and unlocks it on modern edit models

model/conditioning
InstructPixToPixConditioning

Edit an image with a sentence

model/conditioning/instructpix2pix
Invert Mask

One node that flips which half of the image you're fixing

image/mask
Join Audio Channels

The other half of the stereo round-trip

audio
Join Image with Alpha

Turn a mask into a PNG's alpha channel in one node

image/compositing
Extract Text from JSON

Pull one value out of a pile of braces

text
Kandinsky5ImageToVideo

Kandinsky 5's image-to-video node, and the one output everyone misses

model/conditioning/kandinsky
KarrasScheduler

The schedule that ruled SD for years — and why your new model might hate it

model/sampling/schedulers
Kling Avatar 2.0

A broadcast-ready talking avatar from one photo and an audio file

partner/video/Kling
Kling 3.0 First-Last-Frame to Video

The bookend node on the newest model

partner/video/Kling
Kling Image(First Frame) to Video

The classic I2V node, now on turbo

partner/video/Kling
Kling 3.0 Image

The video lab's image node, with reference-image control that's actually useful

partner/image/Kling
Kling 2.6 Image(First Frame) to Video with Audio

Kling 2.6 first-frame to video that ships with sound baked in

partner/video/Kling
Kling Lip Sync Video with Audio

Make a video say what your audio says — Kling lip sync from audio

partner/video/Kling
Kling Lip Sync Video with Text

Make a video say words it never said, no audio file required

partner/video/Kling
Kling Motion Control

Drive a character's movement from a reference video while keeping the look from a still

partner/video/Kling
Kling 3.0 Omni Edit Video

One prompt, keep the sound, swap the world

partner/video/Kling
Kling 3.0 Omni First-Last-Frame to Video

The flexible one that does bookends and multi-reference in a single node

partner/video/Kling
Kling 3.0 Omni Image

The all-in-one Kling node that also edits, and can run a series

partner/image/Kling
Kling 3.0 Omni Image to Video

Seven reference images, storyboards, and audio in one node

partner/video/Kling
Kling 3.0 Omni Text to Video

Storyboards, audio, and a 15-second budget

partner/video/Kling
Kling 3.0 Omni Video to Video

Restyle footage and keep the sound

partner/video/Kling
Kling Start-End Frame to Video

The A-to-B transition node that solves the 'how does it get there' problem

partner/video/Kling
Kling Text to Video

Still the one to learn first

partner/video/Kling
Kling 2.6 Text to Video with Audio

Describe a scene, get sound with it

partner/video/Kling
Kling Video Extend

Your clip is 5 seconds; this makes it longer

partner/video/Kling
Kling 3.0 Video

The one node that does text, image, storyboards, and audio

partner/video/Kling
Krea 2 Image

Krea 2 in one node — no 15GB download, and the style-chaining the open release kept

partner/image/Krea
Krea 2 Style Reference

The chain-node that makes Krea 2 copy a look

partner/image/Krea
KSampler

The one node that's in every workflow you've ever downloaded

model/sampling
KSampler (Advanced)

The same sampler with the training wheels off

model/sampling
KSamplerSelect

Just the sampler dropdown, as an object

model/sampling/samplers
LaplaceScheduler

The sigma curve from a Laplace distribution

model/sampling/schedulers
LatentAdd

Add two latents, get a blend

model/latent/advanced
LatentApplyOperation

Run a latent filter once, before sampling

model/latent/advanced/operations
LatentApplyOperationCFG

Inject a latent filter into every sampling step

model/latent/advanced/operations
Batch Latents (DEPRECATED)

Batch Latents — the deprecated node you should stop wiring in

model/latent/batch
LatentBatchSeedBehavior

Decide whether every item in a batch shares its noise

model/latent/advanced
Latent Blend

Mixing latents the boring, useful way

experimental
Latent Composite

Paste one latent onto another before the sampler

model/latent
Latent Composite Masked

Paste one latent onto another, before sampling

model/latent
LatentConcat

Stitch two latents together — the mechanism behind Kontext-style editing

model/latent/advanced
Crop Latent

Reframe the canvas before the sampler spends a pass

model/latent/transform
LatentCut

Slice a latent like a video clip — frames, strips, and crops

model/latent/advanced
LatentCutToBatch

Slice one latent into a stack of chunks

model/latent/advanced
Flip Latent

Mirror the latent before the model ever sees it

model/latent/transform
Get Latent From Batch

Reach into a stack and pull one out

model/latent/batch
LatentInterpolate

Morph between two images, cleanly

model/latent/advanced
LatentMultiply

The blunt knob for latent brightness

model/latent/advanced
LatentOperationSharpen

Build the sharpen filter, not the sharpened image

model/latent/advanced/operations
LatentOperationTonemapReinhard

Reinhard tonemapping, latent-style

model/latent/advanced/operations
Rotate Latent

Spin your image before it's even an image

model/latent/transform
LatentSubtract

Difference arithmetic on images that aren't images yet

model/latent/advanced
Upscale Latent

The cheap resize that powers hi-res fix

model/latent
Upscale Latent By

Upscaling the latent before you resample

model/latent
Load Latent Upscale Model

Hunyuan Video's in-latent resolution trick

model/loaders
Layers From Bounding Boxes

When your layers arrive as a batch

image
LazyCache

LazyCache — EasyCache's dumber, more compatible sibling

advanced/debug
Load 3D & Animation

The node that put a 3D viewport inside ComfyUI

3d
Load 3D (Advanced)

The mesh loader that doesn't render anything

3d
Load Audio

The front door to every audio workflow in ComfyUI

audio
Load Background Removal Model

Background removal is now a core node — and it's BiRefNet

model/loaders
Load Depth Anything 3

Depth Anything 3 is really a 3D reconstruction model — this is its loader

model/loaders
Load Image

The node almost every workflow starts with

image
Load Image (from Folder)

A whole folder, loaded at once

image
Load Image (as Mask)

Import a mask without the picture

image
Load Image (from Outputs)

Pick up where your last run left off

image
Load Image-Text (from Folder)

Your caption-paired dataset, back as lists

image
Load Latent

Pick up exactly where you left off, in compressed form

model/latent
Load Face Detection Model (MediaPipe)

MediaPipe face detection is now a core node — no face-swap pack required

model/loaders
Load MoGe Model

MoGe gives you the whole geometry, not just a depth map

model/loaders
Load Training Dataset

Skip the re-encode and get straight back to training

model/training
Load Video

Also a core ComfyUI node now — here's how to tell them apart

video
Load Video (from Folder)

Every clip in a folder, in one list

video
Load Video-Text (from Folder)

Videos and their captions, loaded as a pair

video
Load LoRA (Model and CLIP)

The classic loader for SD/SDXL-era LoRAs

model/loaders
Load LoRA (Bypass) (For debugging)

The LoRA loader you'll almost never need

model/loaders
Load LoRA (Bypass, Model Only) (for debugging)

The debugging loader that never touches your weights

model/loaders
Load LoRA

The model-only LoRA loader, and the one most workflows actually use

model/loaders
Load LoRA Model

The training-pipeline LoRA loader with a bypass mode

model/loaders
Extract and Save Lora

Turn a model difference into a LoRA without training

experimental
Plot Loss Graph

The boring node that tells you if training actually worked

model/training
LotusConditioning

The node with no inputs that isn't broken

model/conditioning/lotus
LTX 2.5 Audio To Video

Let the audio track drive the video

partner/video/LTXV
LTX 2.5 Image To Video

Start frame to video on LTX 2.5, cloud-style

partner/video/LTXV
LTX 2.5 Text To Video

LTX 2.5 text to video, without the VRAM dance

partner/video/LTXV
Load LTXV Audio Text Encoder

The 12B Gemma that makes LTX-2 understand you

model/loaders
LTXVAddGuide

The hidden keyframe node that gives LTX first-frame, last-frame, and mid-video control

model/conditioning/ltxv
LTXV Image To Video

LTXV image-to-video, the fast draft model — as a hosted API

partner/video/LTXV
LTXV Text To Video

Decent, but it's the deprecated one now

partner/video/LTXV
LTXV Audio VAE Decode

The moment your audio becomes hearable

model/latent/ltxv
LTXV Audio VAE Encode

Turn a voice or song into conditioning

model/latent/ltxv
Load LTXV Audio VAE

The decoder half of LTX-2's talking videos

model/loaders
Concat AV Latent

Stitch the audio stream onto your video latent

model/latent/ltxv
LTXVConditioning

The frame rate stamp that keeps LTXV clips honest

model/conditioning/ltxv
LTXV Context Windows

The fast model gets long clips, still without the VRAM bill

model/patch
LTXVCropGuides

The cleanup node that keeps your keyframes out of the final video

model/conditioning/ltxv
LTXV Dual CFG Guider

One CFG for the picture, one for the soundtrack — LTXV's split-personality guider

model/sampling/guiders
LTXV Duration Predictor

Stop guessing frame counts — let the model time your shot

conditioning/video_models
LTXV Empty Latent Audio

The blank tape for LTX-2's sound

model/latent/ltxv
LTXVImgToVideo

The fastest image-to-video starter in the open-weights world

model/conditioning/ltxv
LTXVImgToVideoInplace

LTXVImgToVideoInplace

model/conditioning/ltxv
LTXVLatentUpsampler

The 2x latent upscale that makes LTX-2 quality sane

model/latent/ltxv
LTXV Modality Guidance (A/V coupling)

The dial that makes LTX mouths actually match the words

advanced/guidance
LTXV Preprocess

Feed LTX a Worse Image on Purpose (Yes, Really)

video/preprocessors
LTXV Reference Audio (ID-LoRA)

Clone a voice into your LTX video with LTX Reference Audio (ID-LoRA)

model/conditioning/ltxv
LTXVScheduler

The LTX-specific sigmas the official workflows use

model/sampling/schedulers
Separate AV Latent

Your LTX-2 latent is secretly two latents — split them before you decode

model/latent/ltxv
LTXV Spatio-Temporal Guidance (STG)

The detail dial that keeps LTX video from going soft

advanced/guidance
Luma Concepts

Camera direction as dropdowns instead of prompt gambling

partner/video/Luma
Luma UNI-1 Image Edit

Prompt-edit a photo with Luma's flagship model

partner/image/Luma
Luma Image to Image

Luma img2img with a single dial that decides how much the photo changes

partner/image/Luma
Luma Text to Image

The older Luma node that introduced Comfy to Photon

partner/image/Luma
Luma UNI-1 Image

Luma's current text-to-image, stripped to what matters

partner/image/Luma
Luma Image to Video

First frame, last frame, or both — you choose

partner/video/Luma
Luma Ray 3.2 Extend Video

Keep a clip going forward, or prepend a lead-in

partner/video/Luma
Luma Ray 3.2 Image to Video

5 seconds, anchored at both ends

partner/video/Luma
Luma Ray 3.2 Keyframe

Pin guide images to moments on your timeline

partner/video/Luma
Luma Ray 3.2 Keyframes to Video

The closest thing to an animatic in a prompt

partner/video/Luma
Luma Ray 3.2 Text to Video

The 10-second take Luma won't give its image node

partner/video/Luma
Luma Ray 3.2 Video Edit

Re-render your footage, keep the motion

partner/video/Luma
Luma Ray 3.2 Video Reframe

Change the aspect ratio, let AI fill the gaps

partner/video/Luma
Luma Reference

The boring node that makes Luma's image nodes smart

partner/image/Luma
Luma Text to Video

The dependable prompt-to-clip node

partner/video/Luma
Magnific Image Relight

Reshoot the lighting without reshooting

partner/image/Magnific
Magnific Image Skin Enhancer

Add skin texture back in one call

partner/image/Magnific
Magnific Image Style Transfer

Paint your photo in someone else's look

partner/image/Magnific
Magnific Image Upscale (Creative)

The prompt-driven upscaler that invents detail

partner/image/Magnific
Magnific Image Upscale (Precise V2)

High-fidelity upscaling with sharpness you actually control

partner/image/Magnific
Positive-Biased Guidance

The positive-biased guidance node nobody can explain

experimental
Make Training Dataset

The node that turns images and captions into something trainable

model/training
ManualSigmas

Type your noise schedule by hand — total control, zero guardrails

model/sampling/sigmas
Combine Masks

Your mask-building workbench, with an offset

image/mask
Preview Mask

Preview Mask

image/mask
Convert Mask to Image

The bridge for nodes that refuse to talk to masks

image/mask
Detect Face Landmarks (MediaPipe)

Commercially-clean face detection and 478-point meshes in core

image/detection
Draw Face Mask (MediaPipe)

Face, lips, or eye masks from landmarks — the detail pass starts here

image/detection
Visualize Face Landmarks (MediaPipe)

The face-mesh wireframe overlay, now in core

image/detection
Merge Image Lists (DEPRECATED)

The concat node with a retirement plan

image/batch
Merge Splats

The densify trick that makes sparse splats mesh properly

3d/splat
Merge Text Lists (DEPRECATED)

Deprecated, and probably not what you think it is

text
Meshy: Animate Model

Animate Model — one integer to make your rigged character walk

partner/3d/Meshy
Meshy: Image to Model

Image to Model — one image in, textured GLB and FBX out, credits on the bill

partner/3d/Meshy
Meshy: Multi-Image to Model

Multi-Image to Model — 2 to 4 shots, a cloud mesh, and a texture phase you can skip

partner/3d/Meshy
Meshy: Refine Draft Model

Refine Draft Model — the do-over pass that rescues a lumpy first draft

partner/3d/Meshy
Meshy: Rig Model

Rig Model — the bridge between a generated character and something that can walk

partner/3d/Meshy
Meshy: Text to Model

Text to Model — words in, mesh out, and the few prompts that actually work

partner/3d/Meshy
Meshy: Texture Model

Texture Model — re-skinning an existing mesh with text or a reference image

partner/3d/Meshy
MiniMax H3 Image to Video

Prompt in, video-plus-audio latent out

model/conditioning/minimax
MiniMax H3 Reference to Video

Point MiniMax H3 at a face, a clip, and a voice — then talk about them by tag

model/conditioning/minimax
ModelSamplingMiniMaxH3

MiniMax H3's Two Shift Dials — Video and Audio, Tuned Together

model/patch/minimax
MiniMax H3 Context IR (Prompt Enhancer)
partner/video/MiniMax
MiniMax H3 First-Last-Frame to Video

H3 first-last-frame

partner/video/MiniMax
MiniMax H3 Reference to Video

H3 with reference images, video and audio

partner/video/MiniMax
MiniMax H3 Regenerate to 2K
partner/video/MiniMax
MiniMax H3 Text to Video

Text to video, up to 2K

partner/video/MiniMax
MiniMax Hailuo 02 Video

The text-to-video node with a deceptively long input list

partner/video/MiniMax
MiniMax Image to Video

The first-frame node that's more flexible than the flat price suggests

partner/video/MiniMax
MiniMax Music3 Text Encode
model/conditioning/minimax music
MiniMax Text to Video

The plainest API node in the whole category, and why that's a feature

partner/video/MiniMax
ModelAttentionBackend

ModelAttentionBackend — Pick an Attention Engine Per Model, No Restart

model/patch
ModelComputeDtype

ModelComputeDtype — the precision override for when a model comes out wrong

advanced/debug
ModelMergeAdd

The pure-addition merge for stacking models (and why you usually shouldn't)

model/merging
ModelMergeAuraflow

The node for the 6.8B Apache-2.0 flow model Pony V7 quietly stood on

model/merging/model specific
ModelMergeBlocks

The input/middle/out merge that makes your first checkpoint

model/merging
ModelMergeCosmos14B

The 36-block sibling in NVIDIA's world-model merge family

model/merging/model specific
ModelMergeCosmos7B

Block-merging NVIDIA's world model, for people who actually run Cosmos

model/merging/model specific
ModelMergeCosmosPredict2_14B

NVIDIA's robotics world model, merged for nobody

model/merging/model specific
ModelMergeCosmosPredict2_2B

Merge the robotics backbone that accidentally powers Anima

model/merging/model specific
ModelMergeFlux1

Merging Flux checkpoints, block by block

model/merging/model specific
ModelMergeKrea2

Block-merge the hottest base of 2026, and maybe fix its refusals

model/merging/model specific
ModelMergeLTXV

Technically yes, practically eh

model/merging/model specific
ModelMergeMochiPreview

Block-merging Genmo's 48-block video DiT, preview edition

model/merging/model specific
ModelMergeQwenImage

Block-merging Qwen-Image's 60-block MMDiT, for the Apache-2.0 faithful

model/merging/model specific
ModelMergeSD1

The 30-slider workhorse that merging was built on

model/merging/model specific
ModelMergeSD2

Same node as SD1, because SD2 has the same skeleton

model/merging/model specific
ModelMergeSD3_2B

Block-merge control for the SD3 MMDiT, for the few who still merge it

model/merging/model specific
ModelMergeSD35_Large

Merging an 8B MMDiT that almost nobody actually merges

model/merging/model specific
ModelMergeSDXL

The block-merge node that made the SDXL merge era

model/merging/model specific
ModelMergeSimple

The two-checkpoint blend that starts every merge habit

model/merging
ModelMergeSubtract

The difference merge that removes styles and extracts LoRAs

model/merging
ModelMergeWAN2_1

The Wan 2.1 merge node, block-count warning included

model/merging/model specific
ModelNoiseScale

The tiny patch that tells pixel-space models how noisy they were trained

model/patch
Load Model Patch

The experimental front door to the newest control tech

model/loaders
ModelSamplingAuraFlow

The one knob your Z-Image workflow is probably missing

model/patch
ModelSamplingContinuousEDM

Swap the whole noise model to EDM, v-pred, or flow

model/patch
ModelSamplingContinuousV

The v-prediction knob you'll only touch when a workflow tells you to

model/patch
ModelSamplingDiscrete

Force any checkpoint to speak a different noise language

model/patch
ModelSamplingFlux

The node that keeps Flux working when you leave 1024×1024

model/patch/flux
ModelSamplingLTXV

LTX Video's sampling autopilot — wire in your latent and the shift takes care of itself

model/patch/ltxv
ModelSamplingSD3

Retune the schedule when your SD3.5 image looks off

model/patch/stable diffusion
ModelSamplingStableCascade

A shift knob for a model most people have already buried

model/patch/stable cascade
ModelSave

Export that merged model to a file (LoRAs baked in for free)

model/merging
Run MoGe Inference

The point map node that turns one photo into actual 3D geometry

image/geometry estimation
Run MoGe Panorama Inference

MoGe Panorama Inference stitches depth so you don't have to

image/geometry estimation
Convert MoGe Point Map to Mesh

Turning MoGe geometry into a textured GLB

image/geometry estimation
Render MoGe Geometry

Depth and normal previews, with the DirectX/OpenGL trap explained

image/geometry estimation
Apply Morphology

The erode-and-dilate toolbox that fixes your masks

image/filters
MultiGPU CFG Split

The built-in way to make a second GPU earn its keep

advanced/multigpu
Normalized Attention Guidance

Bring negative prompts back on Turbo, Schnell, and every CFG-1 model

advanced/guidance
Normalize Image Colors

A training-tool that reads like a photo filter — don't be fooled

image/color
NormalizeVideoLatentStart

Fix that flash at the start of your videos with NormalizeVideoLatentStart

model/conditioning
OpenAI ChatGPT Advanced Options

Truncation, token budget, and instructions

partner/text/OpenAI
OpenAI ChatGPT

A real LLM node that reads images and files

partner/text/OpenAI
OpenAI DALL·E 2

The museum piece that still takes a picture

partner/image/OpenAI
OpenAI DALL·E 3

The classic OpenAI generator, still in the graph

partner/image/OpenAI
OpenAI GPT Image 2

The current image model behind an oddly-named node

partner/image/OpenAI
OpenAI GPT Image 2

OpenAI's current generator, with editing built into the same node

partner/image/OpenAI
OpenAI ChatGPT Input Files

Turning your docs and PDFs into LLM context

partner/text/OpenAI
OpenAI Sora - Video (DEPRECATED)

The node that exists on a deadline

partner/video/Sora
OpenRouter LLM

OpenRouter turns ComfyUI into a model router

partner/text/OpenRouter
Load Optical Flow Model

RAFT optical flow for VOID

model/loaders
OptimalStepsScheduler

Pre-tuned 'optimal' schedules for Flux, Wan, and Chroma — no thinking required

model/sampling/schedulers
Painter

Paint your inpainting mask right in the graph

image
Cond Pair Combine

The boring but necessary glue for dual-prompt pipelines

advanced/hooks/cond pair
Cond Pair Set Default Combine

The node that fixes the 'stitched together' look

advanced/hooks/cond pair
Cond Pair Set Props

Stamp hooks, masks, and timing onto both your conds at once

advanced/hooks/cond pair
Cond Pair Set Props Combine

Two LoRAs in one image, each locked to its own side

advanced/hooks/cond pair
PatchModelAddDownscale (Kohya Deep Shrink)

The SDXL composition trick that survived

model/patch/unet
Perp-Neg (DEPRECATED by Perp-Neg Guider)

The deprecated perp-neg node, and why you should use the Guider instead

experimental
Perp-Neg Guider

Let your negative prompt stop fighting your positive one

experimental
PerturbedAttentionGuidance

The CFG-style guidance that works when you have no negative prompt

model/patch/unet
PhotoMaker Encode

Put a real face into SD1.5/SDXL prompts without training a LoRA

model/conditioning/photomaker
Load PhotoMaker Model

Identity from a few reference photos, before InstantID made it cool

model/loaders
PiD Conditioning

Attach the latent so the pixel decoder knows what it's decoding

model/conditioning
PixVerse Image to Video

Animate a still on a hosted tier

partner/video/PixVerse
PixVerse Template

Preset styles you bolt onto the video nodes

partner/video/PixVerse
PixVerse Text to Video

Prompt-only clips with a motion-mode switch

partner/video/PixVerse
PixVerse Transition Video

Morph between two images in one shot

partner/video/PixVerse
PolyexponentialScheduler

Karras's flexible cousin

model/sampling/schedulers
Porter-Duff Image Composite

The compositing math node you'll use once a year — and then be glad it exists

image/compositing
Preview 3D & Animation

The zero-effort way to actually see your model

3d
Preview 3D (Advanced)

The preview that passes the mesh through

3d
Preview as Text

The print() statement ComfyUI never had — preview anything as text

utilities
Preview Audio

Hear your track before you commit to a file

audio
Preview Splat

Gaussian splats finally have a first-class viewer

3d
Preview Image

The save node that deliberately doesn't save

image
Preview Point Cloud

The node for raw dots, when you need raw dots

3d
Boolean

The Boolean node is a checkbox. That's more useful than it sounds.

utilities/primitive
Bounding Box

Drag a crop region instead of typing four numbers

utilities/primitive
Float

The Float node isn't a number box — it's how you stop retyping cfg and denoise

utilities/primitive
Int

The node that turned the XYZ plot into a batch queue

utilities/primitive
Text

The deprecated prompt box, and the better node to grab instead

utilities/primitive
Text (Multiline)

The node beginners assume is a text encoder (it isn't)

utilities/primitive
Load CLIP (Quadruple)

Four text encoders in one CLIP output — the HiDream I1 special

model/loaders
Quiver Image to SVG

Turn a raster image into real vector art

partner/image/Quiver
Quiver Text to SVG

Type a prompt, get editable vector graphics

partner/image/Quiver
Apply Qwen Image DiffSynth ControlNet

Structure control for Qwen-Image's editing power

model/patch/qwen
Qwen Image 3 Edit

Edit up to three images with a sentence

partner/image/Qwen
Qwen Image 3 Text to Image

Qwen-Image 3.0 without the 20B model on your disk

partner/image/Qwen
Crop Image (Random)

Same shape, different middle, every run

image/transform
RandomNoise

The seed, as an object you can hand to a sampler

model/sampling/noise
Rebatch Images

How to process 100 frames without nuking your VRAM

image/batch
Rebatch Latents

Rebatch latents when the batch size stops fitting your plan

model/latent/batch
Record Audio

A microphone, built into the graph

audio
Recraft Color RGB

The local helper node that builds a color for Recraft

partner/image/Recraft
Recraft Controls

The free local node that wires palettes into Recraft generation

partner/image/Recraft
Recraft Create Style

Turn a few reference images into a style you can reuse

partner/image/Recraft
Recraft Creative Upscale Image

Creative upscale, with a comma. This one invents detail

partner/image/Recraft
Recraft Crisp Upscale Image

The one-input upscaler that makes things bigger without making things weird

partner/image/Recraft
Recraft Image Inpainting

The mask still wins when the rest of the image has to stay put

partner/image/Recraft
Recraft V3 Image to Image

Image to image with a strength dial you'll actually feel

partner/image/Recraft
Recraft Remove Background

Cut out the background, get the mask for free

partner/image/Recraft
Recraft Replace Background

Swap the backdrop with a sentence, keep the subject

partner/image/Recraft
Recraft Style - Digital Illustration

Forty illustration substyles, one little node

partner/image/Recraft
Recraft Style - Infinite Style Library

The niche node for brand consistency

partner/image/Recraft
Recraft Style - Logo Raster

The logo style node that refuses to say 'none'

partner/image/Recraft
Recraft Style - Realistic Image

The style node you didn't know was already defaulting

partner/image/Recraft
Recraft V3 Text to Image

The 'red panda' model, now a node

partner/image/Recraft
Recraft V3 Text to Vector

Real SVG

partner/image/Recraft
Recraft V4 Text to Image

Recraft's current flagship, and the negative prompt is a lie

partner/image/Recraft
Recraft V4 Text to Vector

The node that outputs actual vector files

partner/image/Recraft
Recraft Vectorize Image

Raster in, editable SVG out, zero tracing pain

partner/image/Recraft
Set Reference Latent

The tiny node that feeds edit models their target

model/conditioning
Set Reference Audio

Reference audio for music generation

model/conditioning
Extract Text

The regex scraper that turns one string into another

text
Match Text

A regex test that gives you a clean boolean

text
Replace Text (Regex)

When plain find-and-replace isn't enough

text
Remove Background

Cut any subject out in one step — and it ships with ComfyUI

image/background removal
Render Splat

The node that turns an invisible cloud of gaussians into an image

3d/splat
RenormCFG

The fix for Lumina 2's bloated, oversaturated high-CFG output

model/patch
Repeat Image Batch

The copy node that's easy to misread

image/batch
Repeat Latent Batch

When you need N copies of the same latent

model/latent/batch
Replace Text (DEPRECATED)

The old dataset-node version of find-and-replace

text
Replace Video Latent Frames

Splice new frames into a video latent like a timeline edit

model/latent/batch
RescaleCFG

The antidote to oversaturated high-CFG renders

model/patch
Resize And Pad Image

Shove any image into a box without wrecking its shape

image/transform
Resize Image/Mask

The one resize node to rule them all (and keep mask and image in sync)

image/transform
Resize Images by Longer Edge (DEPRECATED)

Resize Images by Longer Edge — still in your workflow, but it's been replaced

image/transform
Resize Images by Shorter Edge (DEPRECATED)

Resize Images by Shorter Edge — the deprecated floor-setter

image/transform
Resolution Bucket

Train on mixed image sizes instead of one rigid resolution

model/training
Resolution Selector

Stop doing Empty Latent math by hand — let it work the megapixels for you

utilities
Reve Image Create

Reve, the mystery leaderboard model, as a node

partner/image/Reve
Reve Image Edit

Tell it what to change, in plain English

partner/image/Reve
Reve Image Remix

Mash up to six reference images into something new

partner/image/Reve
Rodin 3D Generate - Detail Generate

Rodin Detail Generate — the older, pricier-feeling sibling that's really just a tier name

partner/3d/Rodin
Rodin 3D Generate - Gen-2 Generate

Rodin Gen-2 is the image-to-3D workhorse — but it runs on their server, not yours

partner/3d/Rodin
Rodin 3D Gen-2.5 - Image to 3D

Rodin Gen-2.5 — the image-to-3D node with a quality dial you can actually afford to read

partner/3d/Rodin
Rodin 3D Gen-2.5 - Text to 3D

Rodin Gen-2.5 Text to 3D — words into a 2M-face model, if your prompt is worth it

partner/3d/Rodin
Rodin 3D Generate - Regular Generate

Rodin Regular Generate — the no-surprises image-to-3D node, before you graduate to Gen-2

partner/3d/Rodin
Rodin 3D Generate - Sketch Generate

Rodin Sketch Generate — the one with just one input, for wireframe-style drafts

partner/3d/Rodin
Rodin 3D Generate - Smooth Generate

Rodin Smooth Generate — for when you want a clean, organic surface, not a voxel statue

partner/3d/Rodin
Run Real-Time Detection (RT-DETR)

Find people, cars and teddy bears in a frame — 80 COCO classes, no text needed

image/detection
Runway Aleph2 Keyframe

Steer your edit at exact moments of the footage

partner/video/Runway
Runway Aleph2 Prompt Image

Pin a reference frame anywhere in Runway's edited video

partner/video/Runway
Runway Aleph2 Video to Video

Edit footage by describing it, and keep the original motion

partner/video/Runway
Runway First-Last-Frame to Video

Runway's first-last-frame trick

partner/video/Runway
Runway Image to Video (Gen3a Turbo)

Image-to-video, the simple way

partner/video/Runway
Runway Image to Video (Gen4 Turbo)

Turn one frame into a clip, fast

partner/video/Runway
Runway Text to Image

Runway's Gen 4 image model, minus the studio app

partner/image/Runway
SAM3 Detect

Text-prompt segmentation and detection in a single core node

image/detection
SAM3 Track Preview

Watch what SAM3 actually tracked before you trust it

image/detection
SAM3 Track to Mask

Pull real masks out of a SAM3 video track

image/detection
Run SAM3 Video Track

Follow one object through every frame of a video

image/detection
Sampler AR Video

The sampler that un-spools video one block at a time

model/sampling/samplers
SamplerCustom

The sampler that sits between 'just click it' and 'build it from parts'

model/sampling/custom
SamplerCustomAdvanced

The assembly point of the sampler graph

model/sampling/custom
SamplerDPMAdaptative

The sampler that decides its own step count — and rarely gets used

model/sampling/samplers
SamplerDPMPP_2M_SDE

The SDE-flavored version of the sampler that ruled SD

model/sampling/samplers
SamplerDPMPP_2S_Ancestral

The fast, never-settling sampler that dominated the SD 1.5 era

model/sampling/samplers
SamplerDPMPP_3M_SDE

The higher-order SDE that squeezes more out of each step

model/sampling/samplers
SamplerDPMPP_SDE

The plain SDE — where DPM++ meets full stochastic sampling

model/sampling/samplers
SamplerER_SDE

The high-order solver that quietly became a default

model/sampling/samplers
SamplerEulerAncestral

The sampler that never settles down — and that's the point

model/sampling/samplers
SamplerEulerAncestralCFG++

The CFG++ sampler for the sub-1 crowd

model/sampling/samplers
SamplerEulerCFG++

The Euler sampler built for the CFG-1 era

experimental
SamplerLCM

LCM's custom sampler, now also a per-step noise dial

model/sampling/samplers
SamplerLCMUpscale

Upscale while you sample, not after

model/sampling/samplers
SamplerLMS

The old default that lost its crown — and why

model/sampling/samplers
SamplerSASolver

The few-step solver ComfyUI now ships with (and the name clash that came along)

model/sampling/samplers
SamplerSEEDS2

One solver, three samplers in a trench coat

model/sampling/samplers
SamplingPercentToSigma

What sigma is 70% of the way through sampling, anyway?

model/sampling/sigmas
Save 3D (Advanced)

The node that finally turns your mesh into a real file

3d
Save Animated PNG

Lossless animation, at a price

image
Save Animated WEBP

The animation format that actually ships

image
Save Audio (FLAC) (DEPRECATED)

Works fine, but it's wearing a DEPRECATED tag for a reason

audio
Save Audio (Advanced)

The one save node you'll actually keep in your graph

audio
Save Audio (MP3) (DEPRECATED)

Lossy, tiny, and replaced by one node

audio
Save Audio (Opus) (DEPRECATED)

The best codec nobody picks, now folded into SaveAudioAdvanced

audio
Save Splat

How to get your gaussian splat out of ComfyUI and into a viewer

3d
Save 3D Model

The export node whose name is a lie (and that's good)

3d
Save Image

The node at the end of every workflow

image
Save Image (Advanced)

The format control the plain node never had

image
Save Image (to Folder) (DEPRECATED)

Deprecated before you found it

image
Save Image-Text (to Folder)

The node that captions your dataset for you

image
Save Latent

Save your work mid-pipeline so you never re-run the slow part

model/latent
Save LoRA Weights

ComfyUI's native 'export my trained LoRA' node, still wearing its experimental badge

model/merging
Save Point Cloud

Writing N×7 clouds to disk as .ply or .npy, depending on who's reading

3d
Save SVG

Vector output for a raster world

image
Save Text

Write any string to disk from the graph

text
Save Training Dataset

Encode once, then train a dozen times without re-encoding

model/training
Save Video

The plain name now points to ComfyUI's own node, not N-Nodes'

video
Save WEBM

The built-in node that writes an actual .webm — alpha channel included

video
Create SCAIL-2 Colored Mask

The color-by-numbers step behind SCAIL-2 character animation

model/conditioning/wan/scail
ScaleROPE

The fix for pushing a video model past its native resolution

model/patch
SD_4XUpscale_Conditioning

The 2023 4x upscaler that's still the cleanest generative upsample

model/conditioning/stable diffusion upscaler
SDPose Draw Keypoints

Turn pose data into the stick figure ControlNet actually wants

image/detection
SDPose Face Bounding Boxes

Face crops from pose data, ready to feed back into the extractor

image/detection
SDPose Keypoint Extractor

Whole-body pose from an SD checkpoint, no external preprocessor

image/detection
SDTurboScheduler

The scheduler that hard-codes Turbo's training ladder

model/sampling/schedulers
Seed

Give your seed a home — one visible value that feeds every sampler

utilities
Apply SeedVR2 Conditioning

The glue that lets Comfy's own sampler drive the best upscaler around

model/conditioning
Post-Process SeedVR2 Output

The node that stops SeedVR2 upscales from going orange

image/post-processors
Pre-Process SeedVR2 Input

The unglamorous node that makes SeedVR2 work

image/pre-processors
Split SeedVR2 Latent

How to make SeedVR2 video upscaling actually fit in VRAM

model/latent/batch
Merge SeedVR2 Latents

Stitch the SeedVR2 chunks back into one video latent

model/latent/batch
Select CLIP Device

Park that huge text encoder on your idle second GPU

advanced/multigpu
Select Model Device

The power node of the multigpu family — and the easiest to misuse

advanced/multigpu
Select VAE Device

The quietest member of the multigpu family, and that's fine

advanced/multigpu
Self-Attention Guidance

Detail without the CFG burn

experimental
Set CLIP Hooks

Apply your LoRA's text-encoder patch to the encoder itself

advanced/hooks/clip
SetFirstSigma

One number at the top of your noise schedule — and why it controls so much

model/sampling/sigmas
Set Hook Keyframes

Make your LoRA fade in, peak, and fall off mid-generation

advanced/hooks/scheduling
Set Latent Noise Mask

The node that actually makes ComfyUI inpaint

model/latent
Set Union ControlNet Type

The dropdown that tells one ControlNet file which of eight modes to run

model/conditioning/controlnet
Shuffle Images List

Make training-style randomness reproducible

image/batch
Shuffle Pairs of Image-Text

Randomize without breaking the pairing

image/batch
Shuffle Videos List

The boring node your video training pipeline quietly needs

video/batch
Shuffle Pairs of Video-Text

Shuffle Your Video-Text Training Pairs Without Fumbling the Captions

dataset/video
SkipLayerGuidanceDiT

Sharper detail without raising CFG

advanced/guidance
SkipLayerGuidanceDiTSimple

The same detail trick at half the price

advanced/guidance
SkipLayerGuidanceSD3

The original SLG, now a compatibility shim

advanced/guidance
Create Solid Mask

A blank canvas of white (or black, or anything in between)

image/mask
Sonilo Text to Music

Prompt to a full track in under a minute, billed by the second

partner/audio/Sonilo
Sonilo Video to Music

Feed it a video, get back a soundtrack that actually matches

partner/audio/Sonilo
Create 3D File (from Splat)

Get your splat off the graph and into a real .ply, .ksplat, or .spz

3d/splat
Extract Mesh from Splat

The node that makes your gaussian splat usable outside ComfyUI

3d/splat
Split Audio Channels

Stereo into two mono tracks

audio
Split Image into List of Tiles

The tiling half of the upscale story

image/batch
Split Image with Alpha

Pull the alpha out of an RGBA image before the graph flattens it to death

image/compositing
SplitSigmas

The splice point for every two-pass workflow

model/sampling/sigmas
SplitSigmasDenoise

Split your denoise budget in two and give each half its own sampler

model/sampling/sigmas
StableCascade_EmptyLatentImage

The two empty latents Stable Cascade needs before you can sample

model/latent/stable cascade
StableCascade_StageB_Conditioning

The bridge between Stable Cascade's two sampling stages

model/conditioning/stable cascade
StableCascade_StageC_VAEEncode

Feed a real image into Stable Cascade's tiny latent space

model/latent/stable cascade
StableCascade_SuperResolutionControlnet

The Stable Cascade super-res node that upsizes nothing (and nobody uses)

experimental/stable cascade
StableZero123_Conditioning

The grandfather of image-to-3D nodes, still holding up

model/conditioning/stable zero123
StableZero123_Conditioning_Batched

Stable Zero123 with batching

model/conditioning/stable zero123
Compare Text

Equal, startswith, endswith as a clean boolean

text
Concatenate Text

The tiny node that glues your whole workflow together

text
Contains Text

One clean boolean for 'is this phrase in there?'

text
Format Text

Format Text

text
Text Length

Count the characters and actually use the number

text
Replace Text

Literal find-and-replace, no regex involved

text
Substring

Carve a slice out of any string

text
Trim Text

Kill the invisible characters before they wreck your prompt

text
Strip Whitespace (DEPRECATED)

The deprecated twin you should quietly retire

text
Apply Style Model

Steal the look without touching the prompt

model/conditioning
Load Style Model

SDXL's reference-style button, quietly still alive

model/loaders
SUPIRApply

Restoration-grade upscaling without a single extension pack

model/patch/supir
SV3D_Conditioning

Turn one photo into an orbiting video with SV3D

model/conditioning/stable video 3d
SVD_img2vid_Conditioning

The old Stable Video Diffusion front door, still in core

model/conditioning/stable video
sync.so Lip Sync

Re-time a face to new audio

partner/video/sync.so
sync.so Talking Image

A still portrait that starts talking

partner/video/sync.so
T5 Tokenizer Options

The quiet tuning knob for T5-encoded models

model/conditioning
Tangential Damping CFG

A zero-knob CFG quality patch you can forget you installed

advanced/guidance
TSR - Temporal Score Rescaling

One knob for diversity, one for when it kicks in

model/patch/unet
Hunyuan3D: 3D Part

3D Part — hand Tencent an FBX and get back a model with labeled, separable parts

partner/3d/Tencent
Hunyuan3D: 3D Texture Edit

3D Texture Edit — repaint a model by describing the new look

partner/3d/Tencent
Hunyuan3D: Image(s) to Model

Image(s) to Model — the node that hands you GLB, OBJ, and the texture maps separately

partner/3d/Tencent
Hunyuan3D: Model to UV

UV unwrapping without the Blender tutorial

partner/3d/Tencent
Hunyuan3D: Smart Topology

From 1.5M triangles to something an engine won't choke on

partner/3d/Tencent
Hunyuan3D: Text to Model

Prompt, face count, and a fully textured GLB

partner/3d/Tencent
TextEncodeAceStepAudio

Prompting ACE-Step for music

model/conditioning/ace
TextEncodeAceStepAudio1.5

TextEncodeAceStepAudio 1.5

model/conditioning/ace
TextEncodeBooguEdit

Instruction editing with the identity preserved

model/conditioning/boogu
TextEncodeHunyuanVideo_ImageToVideo

The Hunyuan I2V encoder, where image and text share one prompt

model/conditioning/hunyuan video
TextEncodeJoyImageEdit

It's called 'text encode', but the image does half the work

model/conditioning/joyimage
TextEncodeMageFlowEdit

Type the edit, point at the image, and let the model do the Photoshopping

model/conditioning/mage
TextEncodeQwenImageEdit

The Qwen-Image-Edit encoder

model/conditioning/qwen image
TextEncodeQwenImageEditPlus

Qwen-Image-Edit-Plus's encoder

model/conditioning/qwen image
TextEncodeZImageOmni

Give Z-Image a reference image without an adapter

model/conditioning/z-image
Generate Text

A real LLM living inside your ComfyUI graph

text
Generate LTX2 Prompt

LTX-2's own prompt enhancer, baked in

text
Draw Text Overlay

The node that stamps a caption on your image without touching the GPU

text
Convert Text to Lowercase (DEPRECATED)

It does what it says, and it's already been replaced

text
Convert Text to Uppercase (DEPRECATED)

Loud, deprecated, still harmless

text
Threshold Mask

Turn a wishy-washy mask into a crisp yes-or-no

image/mask
TomePatchModel

The 2023 speed hack that's mostly retired now

model/patch/unet
Topaz Image Enhance (Legacy)

Topaz 'Reimagine' — the creative upscale that invents the detail you never had

partner/image/Topaz
Topaz Image Enhance

The easy-button upscaler, now inside your graph

partner/image/Topaz
Topaz Video Enhance (Legacy)

Upscaling and slow-mo as an API, and the odd one out in this category

partner/video/Topaz
Topaz Video Enhance

Topaz Video AI, minus the Topaz app

partner/video/Topaz
TorchCompileModel

The fastest way to get 20-30% more speed, with a catch

experimental
Train LoRA

ComfyUI's built-in trainer actually trains — and it's less scary than it looks

model/training
Transform Splat

Move, rotate, and scale a gaussian splat like it's a real object

3d/splat
Trim Audio Duration

The scalpel of the audio family

audio
Trim Video Latent

Cut frames off the front of a video latent — count in latent frames, not pixels

model/latent
Load CLIP (Triple)

The full three-encoder rig for SD3

model/loaders
Tripo: Convert model

Tripo hands you GLB. Blender, Unity and Unreal all want something else.

partner/3d/Tripo
Tripo: Image to Model

Image to Model — a single image, a paid cloud mesh, and the settings that actually change the price

partner/3d/Tripo
Tripo: Import Model

Feed your own 3D models into Tripo's cloud pipeline

partner/3d/Tripo
Tripo: Multiview to Model

Multiview to Model — give it front, left, back, right and stop redrawing the back of your character by hand

partner/3d/Tripo
Tripo P1: Image to Model

Game-ready low-poly from a single image

partner/3d/Tripo
Tripo P1: Multiview to Model

Multiview to Model — Tripo's low-poly specialist, with a mode switch that controls both quality and price

partner/3d/Tripo
Tripo P1: Text to Model

Describe a prop, get a clean low-poly mesh — no image needed

partner/3d/Tripo
Tripo: Refine Draft model

The polish pass only v1.4 Tripo drafts get

partner/3d/Tripo
Tripo: Retarget rigged model

You rigged it. Now make it walk.

partner/3d/Tripo
Tripo: Rig model

One-click auto-rigging for your generated mesh

partner/3d/Tripo
TripoSplat Conditioning

From one photo to a 3D gaussian splat

model/conditioning/triposplat
TripoSplat Preprocess Image

Cut your subject out and park it on a black canvas for TripoSplat

model/conditioning/triposplat
TripoSplat Sampling Preview

Watch your 3D gaussian splat take shape mid-sampling

model/latent/triposplat
Tripo: Text to Model

Type a sentence, get a 3D model — Tripo's flagship node

partner/3d/Tripo
Tripo: Texture model

Your draft is gray plastic. This gives it a paint job.

partner/3d/Tripo
Truncate Text

The 77 that still haunts CLIP

text
Load unCLIP Checkpoint

A fourth output nobody's expecting

model/loaders
unCLIPConditioning

The old-school way to guide a model with a picture

model/conditioning
UNetCrossAttentionMultiply

Scaling your UNet's cross-attention, minus the LoRA

experimental/attention_experiments
Load Diffusion Model

The modern way to load a denoiser by itself

model/loaders
UNetSelfAttentionMultiply

The volume knob on your UNet's self-attention that almost nobody turns

experimental/attention_experiments
UNetTemporalAttentionMultiply

The four-knob attention mixer for video models

experimental/attention_experiments
Load Upscale Model

Load an ESRGAN model for upscaling

model/loaders
Apply USO Style Reference

Reference-image styling for Flux without a LoRA

model/patch/flux
VAE Decode

The only way to actually see what you made

model/latent
VAE Decode Audio

Hear what the latent was saying

model/latent
VAE Decode Audio (Tiled)

Decode long audio latents without blowing up your VRAM

model/latent
VAEDecodeHunyuan3D

Turn a 3D latent into actual voxels you can mesh

model/latent/hunyuan 3d
VAE Decode (Tiled)

Your video and big-image lifeline

model/latent
TripoSplat Decode

From latent to a cloud of 3D gaussians

model/latent/triposplat
VAE Encode

The front door to latent space (every img2img starts here)

model/latent
VAE Encode Audio

The generic door into ComfyUI's audio latent

model/latent
VAE Encode (for Inpainting)

The correct way to start an inpaint in ComfyUI

model/latent
VAE Encode (Tiled)

Get huge images into latent space without OOM

model/latent
Load VAE

The codec node everybody blames and almost nobody understands

model/loaders
VAESave

Save a VAE to a file

model/merging
Google Veo 3 First-Last-Frame to Video

First-frame to last-frame video with Veo 3

partner/video/Veo
Google Veo 3 Video Generation

The one API node that's actually a capability gap

partner/video/Veo
Google Veo 2 Video Generation

Google's Veo 2, without a Google account fight

partner/video/Veo
Sample Video Frame

Cut any video down to exactly N frames

video
Video Linear CFG Guidance

The CFG ramp that keeps early video frames from burning out

model/sampling/guiders
Crop Video (Temporal Random)

Random Temporal Crop Is Your Free Data-Augmentation Trick

video/transform
Trim Video

Cut the dead intro and outro before you pay to process it

video
Crop Video (Temporal)

The Lazy Way to Cut a Video Down to Exactly N Frames

video/transform
Video Triangle CFG Guidance

Dial guidance down at the ends of your video so the middle doesn't melt

model/sampling/guiders
Vidu2 Image-to-Video Generation

Three models, one start frame

partner/video/Vidu
Vidu2 Reference-to-Video Generation

The character-consistency node

partner/video/Vidu
Vidu2 Start/End Frame-to-Video Generation

The bookend that outgrew the 5-second lock

partner/video/Vidu
Vidu2 Text-to-Video Generation

The duration slider finally moves

partner/video/Vidu
Vidu Q3 Image-to-Video Generation

Flagship motion, 2K on the menu

partner/video/Vidu
Vidu Q3 Start/End Frame-to-Video Generation

The expressions, with a guaranteed ending

partner/video/Vidu
Vidu Q3 Text-to-Video Generation

The one people actually talk about

partner/video/Vidu
Vidu Video Extension

Keep the scene going

partner/video/Vidu
Vidu Image To Video Generation

One still, one prompt, a fixed 5 seconds

partner/video/Vidu
Vidu Multi-Frame Video Generation

A storyboard in one generation

partner/video/Vidu
Vidu Reference To Video Generation

Vidu, from up to seven reference images and a prompt

partner/video/Vidu
Vidu Start End To Video Generation

Start and end frames

partner/video/Vidu
Vidu Text To Video Generation

The simplest entry point

partner/video/Vidu
VOIDInpaintConditioning

Inpaint inside a video with VOID's quadmask conditioning

model/conditioning/void
VOID Quadmask Preprocessor

The mask translator for Netflix's video-removal model

image/mask
VOIDSampler

A sampler that exists for exactly one model family — you'll know when you need it

model/sampling/samplers
VOIDWarpedNoise

The trick behind consistent two-pass video

model/latent/void
VOIDWarpedNoiseSource

The adapter that hands VOID's warped noise to SamplerCustomAdvanced

model/latent/void
Voxel to Mesh

Where Hunyuan 3D's blocky world becomes an actual mesh

3d
Voxel to Mesh (Basic) (DEPRECATED)

The deprecated node that still works

3d
VPScheduler

The variance-preserving schedule from the theory books

model/sampling/schedulers
Wan22FunControlToVideo

Reference image plus control video for Wan 2.2

model/conditioning/wan/fun control
Wan22ImageToVideoLatent

The I2V start node that doesn't touch your prompt

model/conditioning/wan
Wan 2.7 Image to Video

Wan 2.7 image-to-video, the version you can't run yourself

partner/video/Wan
Wan 2.7 Reference to Video

The only legal way to run a model Alibaba never released

partner/video/Wan
Wan 2.7 Text to Video

Pure prompt-to-clip on the newest Wan, and the reality check that comes with it

partner/video/Wan
Wan 2.7 Video Continuation

Wan 2.7 continuation, the API way

partner/video/Wan
Wan 2.7 Video Edit

Change the scene without regenerating it

partner/video/Wan
WanAnimate2Cache

Halve your Wan Animate 2 render time for the price of a big stick of RAM

model/conditioning/wan/animate
WanAnimate2ToVideo

Steal the motion from a driving video and put it on your character

model/conditioning/wan/animate
WanAnimateToVideo

Turn a still character into an animated Wan video

model/conditioning/wan/animate
wanBlockSwap

The Wan VRAM saver ComfyUI turned into a no-op

WanCameraEmbedding

Give a Wan video an actual camera move

model/conditioning/wan/camera
WanCameraImageToVideo

Camera-controlled image-to-video

model/conditioning/wan/camera
Wan Context Windows

Push past 81 frames without the identity meltdown

model/patch/wan
WanDancerEncodeAudio

Wan-Dancer's audio analyzer

model/conditioning/wan/dancer
WanDancerPadKeyframes

The node that turns Wan-Dancer keyframes into segments the local model can refine

image/video
WanDancerPadKeyframesList

The same Wan-Dancer keyframe padding, minus the node spaghetti

image/video
WanDancerVideo

Wan-Dancer's conditioning node

model/conditioning/wan/dancer
WanFirstLastFrameToVideo

Two keyframes, an 81-frame movie in between

model/conditioning/wan
WanFunControlToVideo

Wan's control video goes straight into the latent — no ControlNet required

model/conditioning/wan/fun control
WanFunInpaintToVideo

Start and end frames, and the model films the middle

model/conditioning/wan/fun inpaint
WanHuMoImageToVideo

Reference a subject and drive it with audio

model/conditioning/wan/humo
Wan Image to Image

Alibaba's API editor, and the cheapest image node in the partner lineup

partner/image/Wan
WanImageToVideo

The node that turns a still into a Wan video — and where the 81-frame ceiling lives

model/conditioning/wan
Wan Image to Video

The hosted continuation of the open Wan line, now with audio

partner/video/Wan
WanInfiniteTalkToVideo

Make a person talk forever from one photo and one audio file

model/conditioning/wan/infinite talk
WanMoveConcatTrack

Splice two sets of motion tracks into one

model/conditioning/wan/move
WanMoveTracksFromCoords

Paste coordinates, get motion tracks for Wan-Move

model/conditioning/wan/move
WanMoveTrackToVideo

The node that turns drag-points into actual video motion

model/conditioning/wan/move
WanMoveVisualizeTracks

See your motion tracks before you waste a generation on them

model/conditioning/wan/move
WanPhantomSubjectToVideo

Wan's subject-to-video node, with the CFG trick that keeps identity

model/conditioning/wan/phantom subject
Wan Reference to Video

Keep your character (and their voice) with Wan reference-to-video

partner/video/Wan
WanSCAILToVideo

Motion transfer without the stick figure

model/conditioning/wan/scail
WanSoundImageToVideo

Make a character talk, sing, or perform from one image

model/conditioning/wan/sound
WanSoundImageToVideoExtend

Chain S2V clips past the native limit

model/conditioning/wan/sound
Wan Text to Image

ByteDance's Wan image model, hosted so you don't have to be

partner/image/Wan
Wan Text to Video

The Wan that never shipped weights

partner/video/Wan
WanTrackToVideo

Steer a video by dragging points, not by writing a paragraph

model/conditioning/wan/move
Apply Wan Uni3C ControlNet

Apply Wan Uni3C ControlNet — Give Your Wan Camera a Steering Wheel

model/patch/wan
WanVaceToVideo

Reference characters plus a driving video, in one conditioning node

model/conditioning/wan/vace
FlashVSR Video Upscale

Cloud video upscaling without the subscription

partner/video/WaveSpeed
WaveSpeed Image Upscale

WaveSpeed Image Upscale is a hosted upscaler wearing a node costume

partner/image/WaveSpeed
Webcam Capture

Your camera becomes a workflow input

image
Apply Z-Image Fun ControlNet

Pose, depth, and canny for the little model that could

model/patch/z-image
Readme
<div align="center">

ComfyUI

The most powerful and modular AI engine for content creation.

Website Dynamic JSON Badge Twitter Matrix <br>

<!-- Workaround to display total user from https://github.com/badges/shields/issues/4500#issuecomment-2060079995 --> <img width="1590" height="795" alt="ComfyUI Screenshot" src="https://github.com/user-attachments/assets/36e065e0-bfae-4456-8c7f-8369d5ea48a2" /> <br> </div>

ComfyUI is the AI creation engine for visual professionals who demand control over every model, every parameter, and every output. Its powerful and modular node graph interface empowers creatives to generate images, videos, 3D models, audio, and more...

  • ComfyUI natively supports the latest open-source state of the art models.
  • Partner nodes provide access to the best closed source models such as Nano Banana, Seedance, Hunyuan3D, etc.
  • It is available on Windows, Linux, and macOS, locally with our desktop application, our portable install or on our cloud.
  • The most sophisticated workflows can be exposed through a simple UI thanks to App Mode.
  • It integrates seamlessly into production pipelines with our API endpoints.

Get Started

Local

Desktop Application

  • The easiest way to get started.
  • Available on Windows & macOS.

Windows Portable Package

  • Get the latest commits and completely portable.
  • Available on Windows.

Manual Install

Supports all operating systems and GPU types (NVIDIA, AMD, Intel, Apple Silicon, Ascend).

Cloud

Comfy Cloud

  • Our official paid cloud version for those who can't afford local hardware.

Examples

See what ComfyUI can do with the newer template workflows or old example workflows.

Features

  • A visual node graph for building and reusing image, video, audio, 3D, and text workflows without code.
  • Reusable subgraphs, workflow templates, App Mode, and a local API for integrating workflows into applications.
  • Efficient local execution with asynchronous queueing, partial graph re-execution, smart VRAM and RAM management, model offloading, and support for quantized models.
  • Broad native model support. This is a representative list; browse the workflow library for maintained, ready-to-run templates.
    • Image generation: Stable Diffusion 1.5, SDXL, SD3.5, Flux.1, Flux.2, Qwen Image, Z-Image, Hunyuan Image 2.1, HiDream, Lumina Image 2.0, Chroma, Anima, LongCat Image, Ideogram 4, Krea 2, MageFlow, Microsoft Lens, PixelDiT, Kandinsky 5, and Ernie Image.
    • Image editing: Flux Kontext, Flux.2 Klein, Qwen Image Edit, HiDream E1.1 and O1, OmniGen2, Boogu, JoyImage Edit, MageFlow Edit, and LongCat Image Edit.
    • Video generation: Wan 2.1 and 2.2, LTX-Video 2 and 2.3, HunyuanVideo 1.5, Kandinsky 5 Video, CogVideoX, Cosmos Predict2, Bernini-R, SCAIL 2, and Mochi.
    • Audio and video generation: MiniMax H3 and LTX-AV.
    • Audio generation: ACE-Step 1.5, Stable Audio 3 and MiniMax Music 3
    • 3D and vision: Hunyuan3D 2.1, TripoSplat, SeedVR2, SUPIR, Depth Anything 3, MoGe, SAM 3 and 3.1, RT-DETRv4, and BiRefNet.
    • Text generation: Gemma 3 and 4, Qwen3, Qwen3.5, and Qwen3-VL, including multimodal inputs.
  • Load complete checkpoints or separate diffusion models, VAEs, text encoders, LoRAs, ControlNets, adapters, and upscalers from supported model formats.
  • Built-in tools for inpainting, outpainting, reference conditioning, masks and compositing, model merging, upscaling, frame interpolation, segmentation, depth estimation, and media processing.
  • Save and load workflows as JSON, or recover complete workflows and seeds from supported generated media.
  • Runs fully offline: core does not download anything unless you request it. Use --disable-api-nodes to disable the optional paid Comfy API nodes and force all built-in functionality to stay offline.
  • Extend ComfyUI with custom nodes
  • Configure additional model locations with extra_model_paths.yaml.

Release Process

ComfyUI follows a weekly release cycle targeting Monday but this regularly changes because of model releases or large changes to the codebase. There are three interconnected repositories:

  1. ComfyUI Core

    • Releases a new major stable version (e.g., v0.7.0) roughly every 2 weeks.
    • Starting from v0.4.0 patch versions will be used for fixes backported onto the current stable release.
    • Minor versions will be used for releases off the master branch.
    • Patch versions may still be used for releases on the master branch in cases where a backport would not make sense.
    • Commits outside of the stable release tags may be very unstable and break many custom nodes.
    • Serves as the foundation for the desktop release
  2. Comfy Desktop

    • Builds a new release using the latest stable core version
  3. ComfyUI Frontend

    • Every 2+ weeks frontend updates are merged into the core repository
    • Features are frozen for the upcoming core release
    • Development continues for the next release cycle

Shortcuts

| Keybind | Explanation | |------------------------------------|--------------------------------------------------------------------------------------------------------------------| | Ctrl + Enter | Queue up current graph for generation | | Ctrl + Shift + Enter | Queue up current graph as first for generation | | Ctrl + Alt + Enter | Cancel current generation | | Ctrl + Z/Ctrl + Y | Undo/Redo | | Ctrl + S | Save workflow | | Ctrl + O | Load workflow | | Ctrl + A | Select all nodes | | Alt + C | Collapse/uncollapse selected nodes | | Ctrl + M | Mute/unmute selected nodes | | Ctrl + B | Bypass selected nodes (acts like the node was removed from the graph and the wires reconnected through) | | Delete/Backspace | Delete selected nodes | | Ctrl + Backspace | Delete the current graph | | Space | Move the canvas around when held and moving the cursor | | Ctrl/Shift + Click | Add clicked node to selection | | Ctrl + C/Ctrl + V | Copy and paste selected nodes (without maintaining connections to outputs of unselected nodes) | | Ctrl + C/Ctrl + Shift + V | Copy and paste selected nodes (maintaining connections from outputs of unselected nodes to inputs of pasted nodes) | | Shift + Drag | Move multiple selected nodes at the same time | | Ctrl + D | Load default graph | | Alt + + | Canvas Zoom in | | Alt + - | Canvas Zoom out | | Ctrl + Shift + LMB + Vertical drag | Canvas Zoom in/out | | P | Pin/Unpin selected nodes | | Ctrl + G | Group selected nodes | | Q | Toggle visibility of the queue | | H | Toggle visibility of history | | R | Refresh graph | | F | Show/Hide menu | | . | Fit view to selection (Whole graph when nothing is selected) | | Double-Click LMB | Open node quick search palette | | Shift + Drag | Move multiple wires at once | | Ctrl + Alt + LMB | Disconnect all wires from clicked slot |

Ctrl can also be replaced with Cmd instead for macOS users

Installing

Windows and Mac

We highly recommend using the desktop app:

Link to Download

The desktop app is the easiest and best way to use ComfyUI for new users.

Windows Portable

There is a portable standalone build for Windows that should work for running on Nvidia GPUs or for running on your CPU only. It is not recommended for regular users. Regular users should use the desktop app above.

Direct link to download (nvidia)

Simply download, extract with 7-Zip or with the windows explorer on recent windows versions and run. For smaller models you normally only need to put the checkpoints (the huge ckpt/safetensors files) in: ComfyUI\models\checkpoints but many of the larger models have multiple files. Make sure to follow the instructions to know which subfolder to put them in ComfyUI\models\

If you have trouble extracting it, right click the file -> properties -> unblock

The portable above currently comes with python 3.13 and pytorch cuda 13.0. Update your Nvidia drivers if it doesn't start.

All Official Portable Downloads:

Portable for AMD GPUs

Portable for Intel GPUs

Portable for Nvidia GPUs (supports 20 series and above).

Portable for Nvidia GPUs with pytorch cuda 12.6 and python 3.12 (Supports Nvidia 10 series and older GPUs, DO NOT USE THIS ON NEWER 20 SERIES AND ABOVE GPUS).

How do I share models between another UI and ComfyUI?

See the Config file to set the search paths for models. In the standalone windows build you can find this file in the ComfyUI directory. Rename this file to extra_model_paths.yaml and edit it with your favorite text editor.

comfy-cli

You can install and start ComfyUI using comfy-cli:

pip install comfy-cli
comfy install

Manual Install (Windows, Linux)

Python 3.14 works but some custom nodes may have issues. The free threaded variant works but some dependencies will enable the GIL so it's not fully supported.

Python 3.13 is very well supported. If you have trouble with some custom node dependencies on 3.13 you can try 3.12

torch 2.7 is minimally supported but using a newer version is extremely recommended. Using a cu130 or above version of pytorch is required on Nvidia 20 series and above. Some features and optimizations might only work on newer versions. We generally recommend using the latest major version of pytorch with the latest cuda version unless it is less than 2 weeks old. If your pytorch is more than 6 months old, please update it.

Instructions:

Git clone this repo.

Put your SD checkpoints (the huge ckpt/safetensors files) in: models/checkpoints

Put your VAE in: models/vae

AMD GPUs (Linux)

AMD users can install rocm and pytorch with pip if you don't have it already installed, this is the command to install the stable version:

pip install torch torchvision torchaudio --index-url https://download.pytorch.org/whl/rocm7.2

This is the command to install the nightly with ROCm 7.2 which might have some performance improvements:

pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/rocm7.2

AMD GPUs (Experimental: Windows and Linux), RDNA 3, 3.5 and 4 only.

These have less hardware support than the builds above but they work on windows. You also need to install the pytorch version specific to your hardware.

RDNA 3 (RX 7000 series):

pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx110X-all/

RDNA 3.5 (Strix halo/Ryzen AI Max+ 365):

pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx1151/

RDNA 4 (RX 9000 series):

pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx120X-all/

Intel GPUs (Windows and Linux)

Intel Arc GPU users can install native PyTorch with torch.xpu support using pip. More information can be found here

  1. To install PyTorch xpu, use the following command:

pip install torch torchvision torchaudio --index-url https://download.pytorch.org/whl/xpu

This is the command to install the Pytorch xpu nightly which might have some performance improvements:

pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/xpu

NVIDIA

Nvidia users should install stable pytorch using this command:

pip install torch torchvision torchaudio --extra-index-url https://download.pytorch.org/whl/cu130

This is the command to install pytorch nightly instead which might have performance improvements.

pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/cu132

Troubleshooting

If you get the "Torch not compiled with CUDA enabled" error, uninstall torch with:

pip uninstall torch

And install it again with the command above.

Dependencies

Install the dependencies by opening your terminal inside the ComfyUI folder and:

pip install -r requirements.txt

After this you should have everything installed and can proceed to running ComfyUI.

Others:

Apple Mac silicon

You can install ComfyUI in Apple Mac silicon (M1, M2, M3 or M4) with any recent macOS version.

  1. Install pytorch nightly. For instructions, read the Accelerated PyTorch training on Mac Apple Developer guide (make sure to install the latest pytorch nightly).
  2. Follow the ComfyUI manual installation instructions for Windows and Linux.
  3. Install the ComfyUI dependencies. If you have another Stable Diffusion UI you might be able to reuse the dependencies.
  4. Launch ComfyUI by running python main.py

Note: Remember to add your models, VAE, LoRAs etc. to the corresponding Comfy folders, as discussed in ComfyUI manual installation.

Ascend NPUs

For models compatible with Ascend Extension for PyTorch (torch_npu). To get started, ensure your environment meets the prerequisites outlined on the installation page. Here's a step-by-step guide tailored to your platform and installation method:

  1. Begin by installing the recommended or newer kernel version for Linux as specified in the Installation page of torch-npu, if necessary.
  2. Proceed with the installation of Ascend Basekit, which includes the driver, firmware, and CANN, following the instructions provided for your specific platform.
  3. Next, install the necessary packages for torch-npu by adhering to the platform-specific instructions on the Installation page.
  4. Finally, adhere to the ComfyUI manual installation guide for Linux. Once all components are installed, you can run ComfyUI as described earlier.

Cambricon MLUs

For models compatible with Cambricon Extension for PyTorch (torch_mlu). Here's a step-by-step guide tailored to your platform and installation method:

  1. Install the Cambricon CNToolkit by adhering to the platform-specific instructions on the Installation
  2. Next, install the PyTorch(torch_mlu) following the instructions on the Installation
  3. Launch ComfyUI by running python main.py

Iluvatar Corex

For models compatible with Iluvatar Extension for PyTorch. Here's a step-by-step guide tailored to your platform and installation method:

  1. Install the Iluvatar Corex Toolkit by adhering to the platform-specific instructions on the Installation
  2. Launch ComfyUI by running python main.py

ComfyUI-Manager

ComfyUI-Manager is an extension that allows you to easily install, update, and manage custom nodes for ComfyUI.

Setup

  1. Install the manager dependencies:

    pip install -r manager_requirements.txt
    
  2. Enable the manager with the --enable-manager flag when running ComfyUI:

    python main.py --enable-manager
    

Command Line Options

| Flag | Description | |------|-------------| | --enable-manager | Enable ComfyUI-Manager | | --enable-manager-legacy-ui | Use the legacy manager UI instead of the new UI (implies --enable-manager) | | --disable-manager-ui | Disable the manager UI and endpoints while keeping background features like security checks and scheduled installation completion (requires --enable-manager) |

Running

python main.py

For AMD cards not officially supported by ROCm

Try running it with this command if you have issues:

For 6700, 6600 and maybe other RDNA2 or older: HSA_OVERRIDE_GFX_VERSION=10.3.0 python main.py

For AMD 7600 and maybe other RDNA3 cards: HSA_OVERRIDE_GFX_VERSION=11.0.0 python main.py

AMD ROCm Tips

You can try setting this env variable PYTORCH_TUNABLEOP_ENABLED=1 which might speed things up at the cost of a very slow initial run.

Notes

Only parts of the graph that have an output with all the correct inputs will be executed.

Only parts of the graph that change from each execution to the next will be executed, if you submit the same graph twice only the first will be executed. If you change the last part of the graph only the part you changed and the part that depends on it will be executed.

Dragging a generated png on the webpage or loading one will give you the full workflow including seeds that were used to create it.

You can use () to change emphasis of a word or phrase like: (good code:1.2) or (bad code:0.8). The default emphasis for () is 1.1. To use () characters in your actual prompt escape them like \( or \).

You can use {day|night}, for wildcard/dynamic prompts. With this syntax "{wild|card|test}" will be randomly replaced by either "wild", "card" or "test" by the frontend every time you queue the prompt. To use {} characters in your actual prompt escape them like: \{ or \}.

Dynamic prompts also support C-style comments, like // comment or /* comment */.

To use a textual inversion concepts/embeddings in a text prompt put them in the models/embeddings directory and use them in the CLIPTextEncode node like this (you can omit the .pt extension):

embedding:embedding_filename.pt

How to show high-quality previews?

Use --preview-method auto to enable previews.

The default installation includes a fast latent preview method that's low-resolution. To enable higher-quality previews with TAESD, download the taesd_decoder.pth, taesdxl_decoder.pth, taesd3_decoder.pth and taef1_decoder.pth and place them in the models/vae_approx folder. Once they're installed, restart ComfyUI and launch it with --preview-method taesd to enable high-quality previews.

How to use TLS/SSL?

Generate a self-signed certificate (not appropriate for shared/production use) and key by running the command: openssl req -x509 -newkey rsa:4096 -keyout key.pem -out cert.pem -sha256 -days 3650 -nodes -subj "/C=XX/ST=StateName/L=CityName/O=CompanyName/OU=CompanySectionName/CN=CommonNameOrHostname"

Use --tls-keyfile key.pem --tls-certfile cert.pem to enable TLS/SSL, the app will now be accessible with https://... instead of http://....

Note: Windows users can use alexisrolland/docker-openssl or one of the 3rd party binary distributions to run the command example above. <br/><br/>If you use a container, note that the volume mount -v can be a relative path so ... -v ".\:/openssl-certs" ... would create the key & cert files in the current directory of your command prompt or powershell terminal.

Support and dev channel

Discord: Try the #help or #feedback channels.

Matrix space: #comfyui_space:matrix.org (it's like discord but open source).

See also: https://www.comfy.org/

psst — we're hiring! Help build ComfyUI: comfy.org/careers

Frontend Development

As of August 15, 2024, we have transitioned to a new frontend, which is now hosted in a separate repository: ComfyUI Frontend. The compiled JS files (from TS/Vue) are published to pypi and installed as a dependency in ComfyUI.

Reporting Issues and Requesting Features

For any bugs, issues, or feature requests related to the frontend, please use the ComfyUI Frontend repository. This will help us manage and address frontend-specific concerns more efficiently.

Using the Latest Frontend

The new frontend is now the default for ComfyUI. However, please note:

  1. The frontend in the main ComfyUI repository is updated fortnightly.
  2. Daily releases are available in the separate frontend repository.

To use the most up-to-date frontend version:

  1. For the latest daily release, launch ComfyUI with this command line argument:

    --front-end-version Comfy-Org/ComfyUI_frontend@latest
    
  2. For a specific version, replace latest with the desired version number:

    --front-end-version Comfy-Org/[email protected]
    

This approach allows you to easily switch between the stable fortnightly release and the cutting-edge daily updates, or even specific versions for testing purposes.

QA

Which GPU should I buy for this?

See this page for some recommendations