ComfyUI Extension: ComfyUI
ComfyUI is ready to run
It's one of 95 extensions already installed on ComfyICU — nothing to clone, nothing to reconcile. Bring a workflow and you're billed for GPU seconds, not idle time.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
Looking for a different extension?
Custom Nodes (799)
- AddNoise
- Add Text Prefix (DEPRECATED)
- Add Text Suffix (DEPRECATED)
- Adjust Brightness
- Adjust Contrast
- AlignYourStepsScheduler
- Adaptive Projected Guidance
- ARVideoI2V
- Adjust Audio Volume
- Concatenate Audio
- AudioEncoderEncode
- Load Audio Encoder
- Audio Equalizer (3-Band)
- Merge Audio
- Basic Guider
- BasicScheduler
- Batch Images
- Batch Latents
- Batch Masks
- Beeble SwitchX Image Edit
- Beeble SwitchX Video Edit
- Bernini Conditioning
- BetaSamplingScheduler
- Bria FIBO Image Edit
- Bria Remove Image Background
- Bria Remove Video Background
- Bria Remove Video Background (Transparent)
- Bria Video Green Screen
- Bria Video Replace Background
- Build JSON Prompt (Ideogram)
- ByteDance Seedance 2.0 First-Last-Frame to Video
- ByteDance Seedance 2.0 Reference to Video
- ByteDance Seedance 2.0 Text to Video
- ByteDance Create Image Asset
- ByteDance Create Video Asset
- ByteDance First-Last-Frame to Video
- ByteDance Image
- ByteDance Reference Images to Video
- ByteDance Image to Video
- ByteDance Seed
- ByteDance Seedream 4.5 & 5.0
- ByteDance Seedream 4.5 & 5.0
- ByteDance Text to Video
- Detect Edges (Canny)
- Convert Text Case
- Crop Image (Center)
- CFG Guider
- CFGNorm
- CFG Override
- CFGZeroStar
- Load Checkpoint With Config (DEPRECATED)
- Load Checkpoint
- Save Checkpoint
- ChromaRadianceOptions
- Anthropic Claude
- CLIPAttentionMultiply
- Load CLIP
- CLIPMergeAdd
- CLIPMergeSimple
- CLIPMergeSubtract
- CLIPSave
- CLIP Set Last Layer
- CLIP Text Encode (Prompt)
- CLIP Text Encode (Controlnet)
- CLIPTextEncodeFlux
- CLIP Text Encode (HiDream)
- CLIP Text Encode (Hunyuan Image)
- CLIP Text Encode (Kandinsky 5)
- CLIP Text Encode (Lumina 2)
- CLIP Text Encode (PixArt Alpha)
- CLIP Text Encode (SD3)
- CLIP Text Encode (SDXL)
- CLIP Text Encode (SDXL Refiner)
- CLIP Vision Encode
- Load CLIP Vision
- Color Picker
- Transfer Color
- Combine Hooks [2]
- Combine Hooks [4]
- Combine Hooks [8]
- And
- Math Expression
- Not
- Convert Number
- Or
- If/Else Switch
- Conditioning (Average)
- Conditioning (Combine)
- Conditioning (Concat)
- Conditioning (Multiply)
- Conditioning (Set Area)
- Conditioning (Set Area with Percentage)
- Conditioning (Set Area with Percentage for Video)
- Conditioning (Set Area Strength)
- Cond Set Default Combine
- Conditioning (Set Mask)
- Cond Set Props
- Cond Set Props Combine
- ConditioningSetTimestepRange
- ConditioningStableAudio
- Timesteps Range
- Conditioning Zero Out
- Context Windows (Manual)
- Apply ControlNet (DEPRECATED)
- Apply ControlNet
- Apply Controlnet with VAE
- Apply ControlNet Inpainting (AliMama)
- Load ControlNet Model
- Convert Array to String
- Convert Dictionary to String
- CosmosImageToVideoLatent
- CosmosPredict2ImageToVideoLatent
- Create Bounding Boxes
- Create Camera Info
- Create Hook Keyframe
- Create Hook Keyframes From Floats
- Create Hook Keyframes Interp.
- Create Hook LoRA
- Create Hook LoRA (MO)
- Create Hook Model as LoRA
- Create Hook Model as LoRA (MO)
- Create List
- Create Video
- Crop By Bounding Boxes
- Crop Mask
- Curve Editor
- Custom Combo
- Convert DA3 Geometry to Mesh
- Run Depth Anything 3
- Render Depth Anything 3
- Load ControlNet Model (diff)
- Differential Diffusion
- Load Diffusers Model (DEPRECATED)
- DisableNoise
- Draw BBoxes
- Dual CFG Guider
- Load CLIP (Dual)
- Dual Model CFG Guider
- EasyCache
- ElevenLabs Voice Isolation
- ElevenLabs Instant Voice Clone
- ElevenLabs Speech to Speech
- ElevenLabs Speech to Text
- ElevenLabs Text to Dialogue
- ElevenLabs Text to Sound Effects
- ElevenLabs Text to Speech
- ElevenLabs Voice Selector
- Empty Ace Step 1.5 Latent Audio
- Empty Ace Step 1.0 Latent Audio
- EmptyARVideoLatent
- Empty Audio
- EmptyChromaRadianceLatentImage
- EmptyCosmosLatentVideo
- Empty Flux 2 Latent
- Empty HiDream-O1 Latent Image
- EmptyHunyuanImageLatent
- Empty HunyuanVideo 1.0 Latent
- Empty HunyuanVideo 1.5 Latent
- Empty Image
- Empty Latent Audio
- EmptyLatentHunyuan3Dv2
- Empty Latent Image
- EmptyLTXVLatentVideo
- EmptyMochiLatentVideo
- Empty Qwen Image Layered Latent
- EmptySD3LatentImage
- Epsilon Scaling
- ExponentialScheduler
- ExtendIntermediateSigmas
- Feather Mask
- Get Splat
- FlipSigmas
- Flux.2 Image
- Flux.2 [max] Image
- Flux.2 [pro] Image
- Flux2Scheduler
- FluxDisableGuidance
- Flux Erase Image
- FluxGuidance
- FluxKontextImageScale
- Flux.1 Kontext [max] Image
- Edit Model Reference Method
- Flux.1 Kontext [pro] Image
- Flux KV Cache
- Flux.1 Expand Image
- Flux.1 Fill Image
- Flux 1.1 [pro] Ultra Image
- Flux Virtual Try-On
- Run Frame Interpolation Model
- Load Frame Interpolation Model
- FreeU
- FreeU_V2
- FreSca
- Nano Banana Pro (Google Gemini Image)
- Nano Banana (Google Gemini Image)
- Gemini Input Files
- Nano Banana 2
- Nano Banana 2
- Google Gemini
- Google Gemini
- Google Gemini Omni (Video)
- Generate Video Tracks
- Get IC-LoRA Parameters
- Get Image Size
- Get Splat Count
- Get Video Components
- GITSScheduler
- Load GLIGEN Model
- Apply GLIGEN Text Box
- GLSL Shader
- Grok Image Edit
- Grok Image Edit
- Grok Image
- Grok Video Edit
- Grok Video Extend
- Grok Video
- Grok Reference-to-Video
- Grow Mask
- HappyHorse Image to Video
- HappyHorse Reference to Video
- HappyHorse Text to Video
- HappyHorse Video Edit
- HiDream-O1 Patch Seam Smoothing
- HiDream-O1 Reference Images
- HitPaw General Image Enhance
- HitPaw Video Enhance
- Hunyuan3Dv2Conditioning
- Hunyuan3Dv2ConditioningMultiView
- HunyuanImageToVideo
- Hunyuan Latent Refiner
- HunyuanVideo15ImageToVideo
- Hunyuan Video 15 Latent Upscale With Model
- Hunyuan Video 1.5 Super Resolution
- Load Hypernetwork
- HyperTile
- Ideogram 4 Scheduler
- Ideogram V1
- Ideogram V2
- Ideogram V3
- Ideogram V4
- Add Noise to Image
- Batch Images (DEPRECATED)
- Blend Images
- Blur Image
- Convert Image Color to Mask
- Compare Images
- Image Composite Masked
- Crop Image (DEPRECATED)
- Crop Image
- Deduplicate Images
- Flip Image
- Get Image from Batch
- Make Image Grid
- Image Histogram
- Invert Image Colors
- Merge List of Tiles to Image
- Load Checkpoint Image Only (img2vid model)
- ImageOnlyCheckpointSave
- Pad Image for Outpainting
- Quantize Image
- Image RGB to YUV
- Rotate Image
- Upscale Image
- Upscale Image By
- Scale Image to Max Dimension
- Scale Image to Total Pixels
- Sharpen Image
- Stitch Images
- Convert Image to Mask
- Upscale Image (using Model)
- Image YUV to RGB
- InpaintModelConditioning
- InstructPixToPixConditioning
- Invert Mask
- Join Audio Channels
- Join Image with Alpha
- Extract Text from JSON
- Kandinsky5ImageToVideo
- KarrasScheduler
- Kling Avatar 2.0
- Kling Image to Video (Camera Control)
- Kling Camera Controls
- Kling Text to Video (Camera Control)
- Kling Dual Character Video Effects
- Kling 3.0 First-Last-Frame to Video
- Kling Image(First Frame) to Video
- Kling 3.0 Image
- Kling 2.6 Image(First Frame) to Video with Audio
- Kling Lip Sync Video with Audio
- Kling Lip Sync Video with Text
- Kling Motion Control
- Kling 3.0 Omni Edit Video
- Kling 3.0 Omni First-Last-Frame to Video
- Kling 3.0 Omni Image
- Kling 3.0 Omni Image to Video
- Kling 3.0 Omni Text to Video
- Kling 3.0 Omni Video to Video
- Kling Video Effects
- Kling Start-End Frame to Video
- Kling Text to Video
- Kling 2.6 Text to Video with Audio
- Kling Video Extend
- Kling 3.0 Video
- Kling Virtual Try On
- Krea 2 Image
- Krea 2 Style Reference
- KSampler
- KSampler (Advanced)
- KSamplerSelect
- LaplaceScheduler
- LatentAdd
- LatentApplyOperation
- LatentApplyOperationCFG
- Batch Latents (DEPRECATED)
- LatentBatchSeedBehavior
- Latent Blend
- Latent Composite
- Latent Composite Masked
- LatentConcat
- Crop Latent
- LatentCut
- LatentCutToBatch
- Flip Latent
- Get Latent From Batch
- LatentInterpolate
- LatentMultiply
- LatentOperationSharpen
- LatentOperationTonemapReinhard
- Rotate Latent
- LatentSubtract
- Upscale Latent
- Upscale Latent By
- Load Latent Upscale Model
- LazyCache
- Load 3D & Animation
- Load 3D (Advanced)
- Load Audio
- Load Background Removal Model
- Load Depth Anything 3
- Load Image
- Load Image (from Folder)
- Load Image (as Mask)
- Load Image (from Outputs)
- Load Image-Text (from Folder)
- Load Latent
- Load Face Detection Model (MediaPipe)
- Load MoGe Model
- Load Training Dataset
- Load Video
- Load LoRA (Model and CLIP)
- Load LoRA (Bypass) (For debugging)
- Load LoRA (Bypass, Model Only) (for debugging)
- Load LoRA
- Load LoRA Model
- Extract and Save Lora
- Plot Loss Graph
- LotusConditioning
- Load LTXV Audio Text Encoder
- LTXVAddGuide
- LTXV Image To Video
- LTXV Text To Video
- LTXV Audio VAE Decode
- LTXV Audio VAE Encode
- Load LTXV Audio VAE
- LTXVConcatAVLatent
- LTXVConditioning
- LTXV Context Windows
- LTXVCropGuides
- LTXV Empty Latent Audio
- LTXVImgToVideo
- LTXVImgToVideoInplace
- LTXVLatentUpsampler
- LTXV Preprocess
- LTXV Reference Audio (ID-LoRA)
- LTXVScheduler
- LTXVSeparateAVLatent
- Luma Concepts
- Luma UNI-1 Image Edit
- Luma Image to Image
- Luma Text to Image
- Luma UNI-1 Image
- Luma Image to Video
- Luma Ray 3.2 Extend Video
- Luma Ray 3.2 Image to Video
- Luma Ray 3.2 Keyframe
- Luma Ray 3.2 Keyframes to Video
- Luma Ray 3.2 Text to Video
- Luma Ray 3.2 Video Edit
- Luma Ray 3.2 Video Reframe
- Luma Reference
- Luma Text to Video
- Magnific Image Relight
- Magnific Image Skin Enhancer
- Magnific Image Style Transfer
- Magnific Image Upscale (Creative)
- Magnific Image Upscale (Precise V2)
- Positive-Biased Guidance
- Make Training Dataset
- ManualSigmas
- Combine Masks
- Preview Mask
- Convert Mask to Image
- Detect Face Landmarks (MediaPipe)
- Draw Face Mask (MediaPipe)
- Visualize Face Landmarks (MediaPipe)
- Merge Image Lists (DEPRECATED)
- Merge Splats
- Merge Text Lists (DEPRECATED)
- Meshy: Animate Model
- Meshy: Image to Model
- Meshy: Multi-Image to Model
- Meshy: Refine Draft Model
- Meshy: Rig Model
- Meshy: Text to Model
- Meshy: Texture Model
- MiniMax Hailuo Video
- MiniMax Image to Video
- MiniMax Text to Video
- ModelComputeDtype
- ModelMergeAdd
- ModelMergeAuraflow
- ModelMergeBlocks
- ModelMergeCosmos14B
- ModelMergeCosmos7B
- ModelMergeCosmosPredict2_14B
- ModelMergeCosmosPredict2_2B
- ModelMergeFlux1
- ModelMergeKrea2
- ModelMergeLTXV
- ModelMergeMochiPreview
- ModelMergeQwenImage
- ModelMergeSD1
- ModelMergeSD2
- ModelMergeSD3_2B
- ModelMergeSD35_Large
- ModelMergeSDXL
- ModelMergeSimple
- ModelMergeSubtract
- ModelMergeWAN2_1
- ModelNoiseScale
- Load Model Patch
- ModelSamplingAuraFlow
- ModelSamplingContinuousEDM
- ModelSamplingContinuousV
- ModelSamplingDiscrete
- ModelSamplingFlux
- ModelSamplingLTXV
- ModelSamplingSD3
- ModelSamplingStableCascade
- ModelSave
- Run MoGe Inference
- Run MoGe Panorama Inference
- Convert MoGe Point Map to Mesh
- Render MoGe Geometry
- Apply Morphology
- MultiGPU CFG Split
- Normalized Attention Guidance
- Normalize Image Colors
- NormalizeVideoLatentStart
- OpenAI ChatGPT Advanced Options
- OpenAI ChatGPT
- OpenAI DALL·E 2
- OpenAI DALL·E 3
- OpenAI GPT Image 2
- OpenAI GPT Image 2
- OpenAI ChatGPT Input Files
- OpenAI Sora - Video (DEPRECATED)
- OpenRouter LLM
- Load Optical Flow Model
- OptimalStepsScheduler
- Painter
- Cond Pair Combine
- Cond Pair Set Default Combine
- Cond Pair Set Props
- Cond Pair Set Props Combine
- PatchModelAddDownscale (Kohya Deep Shrink)
- Perp-Neg (DEPRECATED by Perp-Neg Guider)
- Perp-Neg Guider
- PerturbedAttentionGuidance
- PhotoMaker Encode
- Load PhotoMaker Model
- PiD Conditioning
- PixVerse Image to Video
- PixVerse Template
- PixVerse Text to Video
- PixVerse Transition Video
- PolyexponentialScheduler
- Porter-Duff Image Composite
- Preview 3D & Animation
- Preview 3D (Advanced)
- Preview as Text
- Preview Audio
- Preview Splat
- Preview Image
- Preview Point Cloud
- Boolean
- Bounding Box
- Float
- Int
- Text String (DEPRECATED)
- Input Text
- Load CLIP (Quadruple)
- Quiver Image to SVG
- Quiver Text to SVG
- Apply Qwen Image DiffSynth ControlNet
- Crop Image (Random)
- RandomNoise
- Rebatch Images
- Rebatch Latents
- Record Audio
- Recraft Color RGB
- Recraft Controls
- Recraft Create Style
- Recraft Creative Upscale Image
- Recraft Crisp Upscale Image
- Recraft Image Inpainting
- Recraft Image to Image
- Recraft Remove Background
- Recraft Replace Background
- Recraft Style - Digital Illustration
- Recraft Style - Infinite Style Library
- Recraft Style - Logo Raster
- Recraft Style - Realistic Image
- Recraft Text to Image
- Recraft Text to Vector
- Recraft V4 Text to Image
- Recraft V4 Text to Vector
- Recraft Vectorize Image
- Set Reference Latent
- Set Reference Audio
- Extract Text
- Match Text
- Replace Text (Regex)
- Remove Background
- Render Splat
- RenormCFG
- Repeat Image Batch
- Repeat Latent Batch
- Replace Text (DEPRECATED)
- Replace Video Latent Frames
- RescaleCFG
- Resize And Pad Image
- Resize Image/Mask
- Resize Images by Longer Edge (DEPRECATED)
- Resize Images by Shorter Edge (DEPRECATED)
- Resolution Bucket
- Resolution Selector
- Reve Image Create
- Reve Image Edit
- Reve Image Remix
- Rodin 3D Generate - Detail Generate
- Rodin 3D Generate - Gen-2 Generate
- Rodin 3D Gen-2.5 - Image to 3D
- Rodin 3D Gen-2.5 - Text to 3D
- Rodin 3D Generate - Regular Generate
- Rodin 3D Generate - Sketch Generate
- Rodin 3D Generate - Smooth Generate
- Run Real-Time Detection (RT-DETR)
- Runway Aleph2 Keyframe
- Runway Aleph2 Prompt Image
- Runway Aleph2 Video to Video
- Runway First-Last-Frame to Video
- Runway Image to Video (Gen3a Turbo)
- Runway Image to Video (Gen4 Turbo)
- Runway Text to Image
- SAM3 Detect
- SAM3 Track Preview
- SAM3 Track to Mask
- Run SAM3 Video Track
- Sampler AR Video
- SamplerCustom
- SamplerCustomAdvanced
- SamplerDPMAdaptative
- SamplerDPMPP_2M_SDE
- SamplerDPMPP_2S_Ancestral
- SamplerDPMPP_3M_SDE
- SamplerDPMPP_SDE
- SamplerER_SDE
- SamplerEulerAncestral
- SamplerEulerAncestralCFG++
- SamplerEulerCFG++
- SamplerLCM
- SamplerLCMUpscale
- SamplerLMS
- SamplerSASolver
- SamplerSEEDS2
- SamplingPercentToSigma
- Save Animated PNG
- Save Animated WEBP
- Save Audio (FLAC) (DEPRECATED)
- Save Audio (Advanced)
- Save Audio (MP3) (DEPRECATED)
- Save Audio (Opus) (DEPRECATED)
- Save 3D Model
- Save Image
- Save Image (Advanced)
- Save Image (to Folder) (DEPRECATED)
- Save Image-Text (to Folder)
- Save Latent
- Save LoRA Weights
- Save SVG
- Save Training Dataset
- Save Video
- Save WEBM
- Create SCAIL-2 Colored Mask
- ScaleROPE
- SD_4XUpscale_Conditioning
- SDPose Draw Keypoints
- SDPose Face Bounding Boxes
- SDPose Keypoint Extractor
- SDTurboScheduler
- Seed
- Select CLIP Device
- Select Model Device
- Select VAE Device
- Self-Attention Guidance
- Set CLIP Hooks
- SetFirstSigma
- Set Hook Keyframes
- Set Latent Noise Mask
- Set Union ControlNet Type
- Shuffle Images List
- Shuffle Pairs of Image-Text
- SkipLayerGuidanceDiT
- SkipLayerGuidanceDiTSimple
- SkipLayerGuidanceSD3
- Create Solid Mask
- Sonilo Text to Music
- Sonilo Video to Music
- Create 3D File (from Splat)
- Extract Mesh from Splat
- Split Audio Channels
- Split Image into List of Tiles
- Split Image with Alpha
- SplitSigmas
- SplitSigmasDenoise
- Stability AI Audio Inpaint
- Stability AI Audio To Audio
- Stability AI Stable Diffusion 3.5 Image
- Stability AI Stable Image Ultra
- Stability AI Text To Audio
- Stability AI Upscale Conservative
- Stability AI Upscale Creative
- Stability AI Upscale Fast
- StableCascade_EmptyLatentImage
- StableCascade_StageB_Conditioning
- StableCascade_StageC_VAEEncode
- StableCascade_SuperResolutionControlnet
- StableZero123_Conditioning
- StableZero123_Conditioning_Batched
- Compare Text
- Concatenate Text
- Contains Text
- Format Text
- Text Length
- Replace Text
- Substring
- Trim Text
- Strip Whitespace (DEPRECATED)
- Apply Style Model
- Load Style Model
- SUPIRApply
- SV3D_Conditioning
- SVD_img2vid_Conditioning
- T5 Tokenizer Options
- Tangential Damping CFG
- TSR - Temporal Score Rescaling
- Hunyuan3D: 3D Part
- Hunyuan3D: 3D Texture Edit
- Hunyuan3D: Image(s) to Model
- Hunyuan3D: Model to UV
- Hunyuan3D: Smart Topology
- Hunyuan3D: Text to Model
- TextEncodeAceStepAudio
- TextEncodeAceStepAudio1.5
- TextEncodeBooguEdit
- TextEncodeHunyuanVideo_ImageToVideo
- TextEncodeQwenImageEdit
- TextEncodeQwenImageEditPlus
- TextEncodeZImageOmni
- Generate Text
- Generate LTX2 Prompt
- Convert Text to Lowercase (DEPRECATED)
- Convert Text to Uppercase (DEPRECATED)
- Threshold Mask
- TomePatchModel
- Topaz Image Enhance
- Topaz Video Enhance (Legacy)
- Topaz Video Enhance
- TorchCompileModel
- Train LoRA
- Transform Splat
- Trim Audio Duration
- Trim Video Latent
- Load CLIP (Triple)
- Tripo: Convert model
- Tripo: Image to Model
- Tripo: Import Model
- Tripo: Multiview to Model
- Tripo P1: Image to Model
- Tripo P1: Multiview to Model
- Tripo P1: Text to Model
- Tripo: Refine Draft model
- Tripo: Retarget rigged model
- Tripo: Rig model
- TripoSplat Conditioning
- TripoSplat Preprocess Image
- TripoSplat Sampling Preview
- Tripo: Text to Model
- Tripo: Texture model
- Truncate Text
- Load unCLIP Checkpoint
- unCLIPConditioning
- UNetCrossAttentionMultiply
- Load Diffusion Model
- UNetSelfAttentionMultiply
- UNetTemporalAttentionMultiply
- Load Upscale Model
- Apply USO Style Reference
- VAE Decode
- VAE Decode Audio
- VAE Decode Audio (Tiled)
- VAEDecodeHunyuan3D
- VAE Decode (Tiled)
- TripoSplat Decode
- VAE Encode
- VAE Encode Audio
- VAE Encode (for Inpainting)
- VAE Encode (Tiled)
- Load VAE
- VAESave
- Google Veo 3 First-Last-Frame to Video
- Google Veo 3 Video Generation
- Google Veo 2 Video Generation
- Video Linear CFG Guidance
- Trim Video
- Video Triangle CFG Guidance
- Vidu2 Image-to-Video Generation
- Vidu2 Reference-to-Video Generation
- Vidu2 Start/End Frame-to-Video Generation
- Vidu2 Text-to-Video Generation
- Vidu Q3 Image-to-Video Generation
- Vidu Q3 Start/End Frame-to-Video Generation
- Vidu Q3 Text-to-Video Generation
- Vidu Video Extension
- Vidu Image To Video Generation
- Vidu Multi-Frame Video Generation
- Vidu Reference To Video Generation
- Vidu Start End To Video Generation
- Vidu Text To Video Generation
- VOIDInpaintConditioning
- VOID Quadmask Preprocessor
- VOIDSampler
- VOIDWarpedNoise
- VOIDWarpedNoiseSource
- Voxel to Mesh
- Voxel to Mesh (Basic) (DEPRECATED)
- VPScheduler
- Wan22FunControlToVideo
- Wan22ImageToVideoLatent
- Wan 2.7 Image to Video
- Wan 2.7 Reference to Video
- Wan 2.7 Text to Video
- Wan 2.7 Video Continuation
- Wan 2.7 Video Edit
- WanAnimateToVideo
- wanBlockSwap
- WanCameraEmbedding
- WanCameraImageToVideo
- Wan Context Windows
- WanDancerEncodeAudio
- WanDancerPadKeyframes
- WanDancerPadKeyframesList
- WanDancerVideo
- WanFirstLastFrameToVideo
- WanFunControlToVideo
- WanFunInpaintToVideo
- WanHuMoImageToVideo
- Wan Image to Image
- WanImageToVideo
- Wan Image to Video
- WanInfiniteTalkToVideo
- WanMoveConcatTrack
- WanMoveTracksFromCoords
- WanMoveTrackToVideo
- WanMoveVisualizeTracks
- WanPhantomSubjectToVideo
- Wan Reference to Video
- WanSCAILToVideo
- WanSoundImageToVideo
- WanSoundImageToVideoExtend
- Wan Text to Image
- Wan Text to Video
- WanTrackToVideo
- WanVaceToVideo
- FlashVSR Video Upscale
- WaveSpeed Image Upscale
- Webcam Capture
- Apply Z-Image Fun ControlNet
README
ComfyUI
The most powerful and modular AI engine for content creation.
<!-- Workaround to display total user from https://github.com/badges/shields/issues/4500#issuecomment-2060079995 --> <img width="1590" height="795" alt="ComfyUI Screenshot" src="https://github.com/user-attachments/assets/36e065e0-bfae-4456-8c7f-8369d5ea48a2" /> <br> </div>ComfyUI is the AI creation engine for visual professionals who demand control over every model, every parameter, and every output. Its powerful and modular node graph interface empowers creatives to generate images, videos, 3D models, audio, and more...
- ComfyUI natively supports the latest open-source state of the art models.
- API nodes provide access to the best closed source models such as Nano Banana, Seedance, Hunyuan3D, etc.
- It is available on Windows, Linux, and macOS, locally with our desktop application, our portable install or on our cloud.
- The most sophisticated workflows can be exposed through a simple UI thanks to App Mode.
- It integrates seamlessly into production pipelines with our API endpoints.
Get Started
Local
Desktop Application
- The easiest way to get started.
- Available on Windows & macOS.
Windows Portable Package
- Get the latest commits and completely portable.
- Available on Windows.
Manual Install
Supports all operating systems and GPU types (NVIDIA, AMD, Intel, Apple Silicon, Ascend).
Cloud
Comfy Cloud
- Our official paid cloud version for those who can't afford local hardware.
Examples
See what ComfyUI can do with the newer template workflows or old example workflows.
Features
- Nodes/graph/flowchart interface to experiment and create complex Stable Diffusion workflows without needing to code anything.
- NOTE: There are many more models supported than the list below, if you want to see what is supported see our templates list inside ComfyUI.
- Image Models
- SD1.x, SD2.x (unCLIP)
- SDXL, SDXL Turbo
- Stable Cascade
- SD3 and SD3.5
- Pixart Alpha and Sigma
- AuraFlow
- HunyuanDiT
- Flux
- Lumina Image 2.0
- HiDream
- Qwen Image
- Hunyuan Image 2.1
- Flux 2
- Z Image
- Ernie Image
- Image Editing Models
- Video Models
- Audio Models
- 3D Models
- Asynchronous Queue system
- Many optimizations: Only re-executes the parts of the workflow that changes between executions.
- Smart memory management: can automatically run large models on GPUs with as low as 1GB vram with smart offloading.
- Works even if you don't have a GPU with:
--cpu(slow) - Can load ckpt and safetensors: All in one checkpoints or standalone diffusion models, VAEs and CLIP models.
- Safe loading of ckpt, pt, pth, etc.. files.
- Embeddings/Textual inversion
- Loras (regular, locon and loha)
- Hypernetworks
- Loading full workflows (with seeds) from generated PNG, WebP and FLAC files.
- Saving/Loading workflows as Json files.
- Nodes interface can be used to create complex workflows like one for Hires fix or much more advanced ones.
- Area Composition
- Inpainting with both regular and inpainting models.
- ControlNet and T2I-Adapter
- Upscale Models (ESRGAN, ESRGAN variants, SwinIR, Swin2SR, etc...)
- GLIGEN
- Model Merging
- LCM models and Loras
- Latent previews with TAESD
- Works fully offline: core will never download anything unless you want to.
- Optional API nodes to use paid models from external providers through the online Comfy API disable with:
--disable-api-nodes - Config file to set the search paths for models.
Workflow examples can be found on the Examples page
Release Process
ComfyUI follows a weekly release cycle targeting Monday but this regularly changes because of model releases or large changes to the codebase. There are three interconnected repositories:
-
- Releases a new major stable version (e.g., v0.7.0) roughly every 2 weeks.
- Starting from v0.4.0 patch versions will be used for fixes backported onto the current stable release.
- Minor versions will be used for releases off the master branch.
- Patch versions may still be used for releases on the master branch in cases where a backport would not make sense.
- Commits outside of the stable release tags may be very unstable and break many custom nodes.
- Serves as the foundation for the desktop release
-
- Builds a new release using the latest stable core version
-
- Every 2+ weeks frontend updates are merged into the core repository
- Features are frozen for the upcoming core release
- Development continues for the next release cycle
Shortcuts
| Keybind | Explanation |
|------------------------------------|--------------------------------------------------------------------------------------------------------------------|
| Ctrl + Enter | Queue up current graph for generation |
| Ctrl + Shift + Enter | Queue up current graph as first for generation |
| Ctrl + Alt + Enter | Cancel current generation |
| Ctrl + Z/Ctrl + Y | Undo/Redo |
| Ctrl + S | Save workflow |
| Ctrl + O | Load workflow |
| Ctrl + A | Select all nodes |
| Alt + C | Collapse/uncollapse selected nodes |
| Ctrl + M | Mute/unmute selected nodes |
| Ctrl + B | Bypass selected nodes (acts like the node was removed from the graph and the wires reconnected through) |
| Delete/Backspace | Delete selected nodes |
| Ctrl + Backspace | Delete the current graph |
| Space | Move the canvas around when held and moving the cursor |
| Ctrl/Shift + Click | Add clicked node to selection |
| Ctrl + C/Ctrl + V | Copy and paste selected nodes (without maintaining connections to outputs of unselected nodes) |
| Ctrl + C/Ctrl + Shift + V | Copy and paste selected nodes (maintaining connections from outputs of unselected nodes to inputs of pasted nodes) |
| Shift + Drag | Move multiple selected nodes at the same time |
| Ctrl + D | Load default graph |
| Alt + + | Canvas Zoom in |
| Alt + - | Canvas Zoom out |
| Ctrl + Shift + LMB + Vertical drag | Canvas Zoom in/out |
| P | Pin/Unpin selected nodes |
| Ctrl + G | Group selected nodes |
| Q | Toggle visibility of the queue |
| H | Toggle visibility of history |
| R | Refresh graph |
| F | Show/Hide menu |
| . | Fit view to selection (Whole graph when nothing is selected) |
| Double-Click LMB | Open node quick search palette |
| Shift + Drag | Move multiple wires at once |
| Ctrl + Alt + LMB | Disconnect all wires from clicked slot |
Ctrl can also be replaced with Cmd instead for macOS users
Installing
Windows Portable
There is a portable standalone build for Windows that should work for running on Nvidia GPUs or for running on your CPU only on the releases page.
Direct link to download
Simply download, extract with 7-Zip or with the windows explorer on recent windows versions and run. For smaller models you normally only need to put the checkpoints (the huge ckpt/safetensors files) in: ComfyUI\models\checkpoints but many of the larger models have multiple files. Make sure to follow the instructions to know which subfolder to put them in ComfyUI\models\
If you have trouble extracting it, right click the file -> properties -> unblock
The portable above currently comes with python 3.13 and pytorch cuda 13.0. Update your Nvidia drivers if it doesn't start.
All Official Portable Downloads:
Portable for Nvidia GPUs (supports 20 series and above).
Portable for Nvidia GPUs with pytorch cuda 12.6 and python 3.12 (Supports Nvidia 10 series and older GPUs).
How do I share models between another UI and ComfyUI?
See the Config file to set the search paths for models. In the standalone windows build you can find this file in the ComfyUI directory. Rename this file to extra_model_paths.yaml and edit it with your favorite text editor.
comfy-cli
You can install and start ComfyUI using comfy-cli:
pip install comfy-cli
comfy install
Manual Install (Windows, Linux)
Python 3.14 works but some custom nodes may have issues. The free threaded variant works but some dependencies will enable the GIL so it's not fully supported.
Python 3.13 is very well supported. If you have trouble with some custom node dependencies on 3.13 you can try 3.12
torch 2.5 is minimally supported but using a newer version is extremely recommended. Some features and optimizations might only work on newer versions. We generally recommend using the latest major version of pytorch with the latest cuda version unless it is less than 2 weeks old. If your pytorch is more than 6 months old, please update it.
Instructions:
Git clone this repo.
Put your SD checkpoints (the huge ckpt/safetensors files) in: models/checkpoints
Put your VAE in: models/vae
AMD GPUs (Linux)
AMD users can install rocm and pytorch with pip if you don't have it already installed, this is the command to install the stable version:
pip install torch torchvision torchaudio --index-url https://download.pytorch.org/whl/rocm7.2
This is the command to install the nightly with ROCm 7.2 which might have some performance improvements:
pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/rocm7.2
AMD GPUs (Experimental: Windows and Linux), RDNA 3, 3.5 and 4 only.
These have less hardware support than the builds above but they work on windows. You also need to install the pytorch version specific to your hardware.
RDNA 3 (RX 7000 series):
pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx110X-all/
RDNA 3.5 (Strix halo/Ryzen AI Max+ 365):
pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx1151/
RDNA 4 (RX 9000 series):
pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx120X-all/
Intel GPUs (Windows and Linux)
Intel Arc GPU users can install native PyTorch with torch.xpu support using pip. More information can be found here
- To install PyTorch xpu, use the following command:
pip install torch torchvision torchaudio --index-url https://download.pytorch.org/whl/xpu
This is the command to install the Pytorch xpu nightly which might have some performance improvements:
pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/xpu
NVIDIA
Nvidia users should install stable pytorch using this command:
pip install torch torchvision torchaudio --extra-index-url https://download.pytorch.org/whl/cu130
This is the command to install pytorch nightly instead which might have performance improvements.
pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/cu132
Troubleshooting
If you get the "Torch not compiled with CUDA enabled" error, uninstall torch with:
pip uninstall torch
And install it again with the command above.
Dependencies
Install the dependencies by opening your terminal inside the ComfyUI folder and:
pip install -r requirements.txt
After this you should have everything installed and can proceed to running ComfyUI.
Others:
Apple Mac silicon
You can install ComfyUI in Apple Mac silicon (M1, M2, M3 or M4) with any recent macOS version.
- Install pytorch nightly. For instructions, read the Accelerated PyTorch training on Mac Apple Developer guide (make sure to install the latest pytorch nightly).
- Follow the ComfyUI manual installation instructions for Windows and Linux.
- Install the ComfyUI dependencies. If you have another Stable Diffusion UI you might be able to reuse the dependencies.
- Launch ComfyUI by running
python main.py
Note: Remember to add your models, VAE, LoRAs etc. to the corresponding Comfy folders, as discussed in ComfyUI manual installation.
Ascend NPUs
For models compatible with Ascend Extension for PyTorch (torch_npu). To get started, ensure your environment meets the prerequisites outlined on the installation page. Here's a step-by-step guide tailored to your platform and installation method:
- Begin by installing the recommended or newer kernel version for Linux as specified in the Installation page of torch-npu, if necessary.
- Proceed with the installation of Ascend Basekit, which includes the driver, firmware, and CANN, following the instructions provided for your specific platform.
- Next, install the necessary packages for torch-npu by adhering to the platform-specific instructions on the Installation page.
- Finally, adhere to the ComfyUI manual installation guide for Linux. Once all components are installed, you can run ComfyUI as described earlier.
Cambricon MLUs
For models compatible with Cambricon Extension for PyTorch (torch_mlu). Here's a step-by-step guide tailored to your platform and installation method:
- Install the Cambricon CNToolkit by adhering to the platform-specific instructions on the Installation
- Next, install the PyTorch(torch_mlu) following the instructions on the Installation
- Launch ComfyUI by running
python main.py
Iluvatar Corex
For models compatible with Iluvatar Extension for PyTorch. Here's a step-by-step guide tailored to your platform and installation method:
- Install the Iluvatar Corex Toolkit by adhering to the platform-specific instructions on the Installation
- Launch ComfyUI by running
python main.py
ComfyUI-Manager
ComfyUI-Manager is an extension that allows you to easily install, update, and manage custom nodes for ComfyUI.
Setup
-
Install the manager dependencies:
pip install -r manager_requirements.txt -
Enable the manager with the
--enable-managerflag when running ComfyUI:python main.py --enable-manager
Command Line Options
| Flag | Description |
|------|-------------|
| --enable-manager | Enable ComfyUI-Manager |
| --enable-manager-legacy-ui | Use the legacy manager UI instead of the new UI (implies --enable-manager) |
| --disable-manager-ui | Disable the manager UI and endpoints while keeping background features like security checks and scheduled installation completion (requires --enable-manager) |
Running
python main.py
For AMD cards not officially supported by ROCm
Try running it with this command if you have issues:
For 6700, 6600 and maybe other RDNA2 or older: HSA_OVERRIDE_GFX_VERSION=10.3.0 python main.py
For AMD 7600 and maybe other RDNA3 cards: HSA_OVERRIDE_GFX_VERSION=11.0.0 python main.py
AMD ROCm Tips
You can try setting this env variable PYTORCH_TUNABLEOP_ENABLED=1 which might speed things up at the cost of a very slow initial run.
Notes
Only parts of the graph that have an output with all the correct inputs will be executed.
Only parts of the graph that change from each execution to the next will be executed, if you submit the same graph twice only the first will be executed. If you change the last part of the graph only the part you changed and the part that depends on it will be executed.
Dragging a generated png on the webpage or loading one will give you the full workflow including seeds that were used to create it.
You can use () to change emphasis of a word or phrase like: (good code:1.2) or (bad code:0.8). The default emphasis for () is 1.1. To use () characters in your actual prompt escape them like \( or \).
You can use {day|night}, for wildcard/dynamic prompts. With this syntax "{wild|card|test}" will be randomly replaced by either "wild", "card" or "test" by the frontend every time you queue the prompt. To use {} characters in your actual prompt escape them like: \{ or \}.
Dynamic prompts also support C-style comments, like // comment or /* comment */.
To use a textual inversion concepts/embeddings in a text prompt put them in the models/embeddings directory and use them in the CLIPTextEncode node like this (you can omit the .pt extension):
embedding:embedding_filename.pt
How to show high-quality previews?
Use --preview-method auto to enable previews.
The default installation includes a fast latent preview method that's low-resolution. To enable higher-quality previews with TAESD, download the taesd_decoder.pth, taesdxl_decoder.pth, taesd3_decoder.pth and taef1_decoder.pth and place them in the models/vae_approx folder. Once they're installed, restart ComfyUI and launch it with --preview-method taesd to enable high-quality previews.
How to use TLS/SSL?
Generate a self-signed certificate (not appropriate for shared/production use) and key by running the command: openssl req -x509 -newkey rsa:4096 -keyout key.pem -out cert.pem -sha256 -days 3650 -nodes -subj "/C=XX/ST=StateName/L=CityName/O=CompanyName/OU=CompanySectionName/CN=CommonNameOrHostname"
Use --tls-keyfile key.pem --tls-certfile cert.pem to enable TLS/SSL, the app will now be accessible with https://... instead of http://....
Note: Windows users can use alexisrolland/docker-openssl or one of the 3rd party binary distributions to run the command example above. <br/><br/>If you use a container, note that the volume mount
-vcan be a relative path so... -v ".\:/openssl-certs" ...would create the key & cert files in the current directory of your command prompt or powershell terminal.
Support and dev channel
Discord: Try the #help or #feedback channels.
Matrix space: #comfyui_space:matrix.org (it's like discord but open source).
See also: https://www.comfy.org/
psst — we're hiring! Help build ComfyUI: comfy.org/careers
Frontend Development
As of August 15, 2024, we have transitioned to a new frontend, which is now hosted in a separate repository: ComfyUI Frontend. The compiled JS files (from TS/Vue) are published to pypi and installed as a dependency in ComfyUI.
Reporting Issues and Requesting Features
For any bugs, issues, or feature requests related to the frontend, please use the ComfyUI Frontend repository. This will help us manage and address frontend-specific concerns more efficiently.
Using the Latest Frontend
The new frontend is now the default for ComfyUI. However, please note:
- The frontend in the main ComfyUI repository is updated fortnightly.
- Daily releases are available in the separate frontend repository.
To use the most up-to-date frontend version:
-
For the latest daily release, launch ComfyUI with this command line argument:
--front-end-version Comfy-Org/ComfyUI_frontend@latest -
For a specific version, replace
latestwith the desired version number:--front-end-version Comfy-Org/[email protected]
This approach allows you to easily switch between the stable fortnightly release and the cutting-edge daily updates, or even specific versions for testing purposes.
QA
Which GPU should I buy for this?
ComfyUI is ready to run
It's one of 95 extensions already installed on ComfyICU — nothing to clone, nothing to reconcile. Bring a workflow and you're billed for GPU seconds, not idle time.