ComfyUI Extension: ComfyUI
ComfyUI is ready to run
It's one of 95 extensions already installed on ComfyICU — nothing to clone, nothing to reconcile. Bring a workflow and you're billed for GPU seconds, not idle time.
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
Looking for a different extension?
Custom Nodes (799)
- Recraft Color RGB
- Save SVG
- SamplerEulerAncestralCFG++
- Rotate Image
- Save LoRA Weights
- Create Hook Keyframe
- Load GLIGEN Model
- CLIPSave
- Upscale Image By
- Load Video
- Beeble SwitchX Image Edit
- ByteDance Seedance 2.0 First-Last-Frame to Video
- ByteDance Seedance 2.0 Reference to Video
- ByteDance Seed
- Load Upscale Model
- Kling 3.0 Omni Image to Video
- Krea 2 Image
- Conditioning (Set Mask)
- Convert Mask to Image
- Magnific Image Relight
- Kling 3.0 Omni First-Last-Frame to Video
- Rodin 3D Gen-2.5 - Text to 3D
- Topaz Video Enhance
- Tripo P1: Multiview to Model
- Load Image (as Mask)
- Convert Dictionary to String
- Run Frame Interpolation Model
- Apply ControlNet (DEPRECATED)
- SkipLayerGuidanceSD3
- OpenAI GPT Image 2
- CLIP Text Encode (SDXL Refiner)
- Voxel to Mesh
- Contains Text
- VAE Decode
- LTXV Audio VAE Encode
- Set Latent Noise Mask
- Latent Composite
- Latent Blend
- Self-Attention Guidance
- CLIP Text Encode (Controlnet)
- Bria FIBO Image Edit
- CLIP Set Last Layer
- LatentInterpolate
- ElevenLabs Instant Voice Clone
- Upscale Latent By
- Threshold Mask
- Convert Image to Mask
- Save Latent
- Split Image with Alpha
- Apply USO Style Reference
- ModelSamplingDiscrete
- SDTurboScheduler
- Empty HunyuanVideo 1.5 Latent
- TomePatchModel
- SamplerCustom
- Save Animated WEBP
- PixVerse Template
- ModelMergeCosmosPredict2_2B
- Create Hook LoRA
- RandomNoise
- DisableNoise
- AddNoise
- SamplerCustomAdvanced
- Resize Images by Shorter Edge (DEPRECATED)
- ModelNoiseScale
- ModelComputeDtype
- Conditioning (Concat)
- Save Audio (MP3) (DEPRECATED)
- Kling Video Effects
- Create Solid Mask
- Detect Edges (Canny)
- Load Style Model
- Conditioning (Set Area)
- Recraft Style - Logo Raster
- SplitSigmas
- SplitSigmasDenoise
- SetFirstSigma
- LatentMultiply
- CFG Guider
- ModelMergeSD1
- Load Optical Flow Model
- Upscale Image
- Save Checkpoint
- Merge Text Lists (DEPRECATED)
- SamplerEulerAncestral
- Flux KV Cache
- Image RGB to YUV
- Merge Image Lists (DEPRECATED)
- SD_4XUpscale_Conditioning
- Merge Audio
- StableCascade_SuperResolutionControlnet
- InstructPixToPixConditioning
- Load CLIP (Triple)
- ModelMergeCosmosPredict2_14B
- ModelMergeKrea2
- Rebatch Latents
- Repeat Latent Batch
- ModelMergeSD35_Large
- ModelMergeLTXV
- Load unCLIP Checkpoint
- Load Image
- ModelMergeBlocks
- Video Triangle CFG Guidance
- Kling 3.0 Omni Text to Video
- Combine Hooks [2]
- Cond Set Default Combine
- SamplerDPMPP_SDE
- RenormCFG
- Extract Text from JSON
- Convert Array to String
- CLIPMergeSimple
- Image Histogram
- VOIDSampler
- TextEncodeQwenImageEditPlus
- ByteDance Seedream 4.5 & 5.0
- ElevenLabs Text to Dialogue
- Convert Image Color to Mask
- Invert Mask
- Google Gemini Omni (Video)
- Join Image with Alpha
- ModelMergeSubtract
- OpenRouter LLM
- Text Length
- LTXVImgToVideoInplace
- OpenAI GPT Image 2
- Adaptive Projected Guidance
- LotusConditioning
- VAEDecodeHunyuan3D
- TextEncodeAceStepAudio
- Replace Text
- wanBlockSwap
- Convert Number
- Empty Image
- Load VAE
- VAE Encode (Tiled)
- Batch Latents (DEPRECATED)
- Get Latent From Batch
- LatentOperationTonemapReinhard
- SamplerLMS
- SamplerDPMPP_2S_Ancestral
- VAE Encode (for Inpainting)
- ManualSigmas
- HyperTile
- ModelSamplingContinuousEDM
- ModelSamplingContinuousV
- ModelSamplingFlux
- Bounding Box
- Repeat Image Batch
- Load ControlNet Model
- LatentAdd
- LatentCutToBatch
- Quantize Image
- Load LoRA Model
- Batch Images
- BetaSamplingScheduler
- FlipSigmas
- HappyHorse Text to Video
- ModelSamplingStableCascade
- Crop Image (DEPRECATED)
- VAE Decode Audio (Tiled)
- LTXV Audio VAE Decode
- WanMoveVisualizeTracks
- Image YUV to RGB
- Crop Latent
- Merge Splats
- HunyuanVideo15ImageToVideo
- Positive-Biased Guidance
- LTXVConditioning
- Create Hook Model as LoRA (MO)
- Combine Hooks [8]
- Empty Latent Image
- Get Video Components
- Sharpen Image
- Load Checkpoint
- Load LoRA
- LatentConcat
- FreeU_V2
- SamplerSASolver
- WanAnimateToVideo
- Format Text
- Compare Text
- Replace Text (DEPRECATED)
- UNetTemporalAttentionMultiply
- Resize Images by Longer Edge (DEPRECATED)
- Normalize Image Colors
- Apply Morphology
- EmptySD3LatentImage
- TextEncodeQwenImageEdit
- Empty Qwen Image Layered Latent
- Conditioning (Set Area Strength)
- LTXVSeparateAVLatent
- WanMoveConcatTrack
- LTXVLatentUpsampler
- Load 3D & Animation
- Preview Point Cloud
- GLSL Shader
- Load LoRA (Bypass) (For debugging)
- Trim Video Latent
- CLIP Text Encode (HiDream)
- TextEncodeZImageOmni
- VAE Decode (Tiled)
- Painter
- StableCascade_StageB_Conditioning
- ModelMergeMochiPreview
- VOID Quadmask Preprocessor
- Concatenate Text
- Load Face Detection Model (MediaPipe)
- Get Splat
- Get Splat Count
- Adjust Brightness
- Perp-Neg Guider
- EasyCache
- Math Expression
- Render Depth Anything 3
- Flux.2 [max] Image
- Build JSON Prompt (Ideogram)
- Deduplicate Images
- PhotoMaker Encode
- StableCascade_EmptyLatentImage
- ModelMergeQwenImage
- Save Audio (Opus) (DEPRECATED)
- Invert Image Colors
- CLIPMergeSubtract
- Perp-Neg (DEPRECATED by Perp-Neg Guider)
- Grok Image Edit
- Kling 3.0 First-Last-Frame to Video
- Run Real-Time Detection (RT-DETR)
- WanDancerVideo
- WanDancerPadKeyframes
- Luma Ray 3.2 Image to Video
- Bria Remove Video Background (Transparent)
- Nano Banana Pro (Google Gemini Image)
- Nano Banana 2
- Nano Banana 2
- Grok Reference-to-Video
- Hunyuan3D: Text to Model
- Hunyuan3D: Image(s) to Model
- LTXV Empty Latent Audio
- ModelMergeAdd
- Kling 3.0 Omni Video to Video
- Luma Image to Video
- Recraft Style - Digital Illustration
- Recraft Create Style
- Rodin 3D Generate - Detail Generate
- Meshy: Refine Draft Model
- OpenAI ChatGPT Input Files
- HappyHorse Image to Video
- Create Bounding Boxes
- Sonilo Text to Music
- Extract and Save Lora
- Stability AI Audio To Audio
- Tripo: Retarget rigged model
- SUPIRApply
- LTXV Reference Audio (ID-LoRA)
- Select VAE Device
- Load Frame Interpolation Model
- Kling Text to Video
- Get Image from Batch
- Set Hook Keyframes
- ByteDance First-Last-Frame to Video
- LTXVConcatAVLatent
- MultiGPU CFG Split
- Preview 3D & Animation
- EmptyARVideoLatent
- ExponentialScheduler
- ByteDance Image to Video
- Draw Face Mask (MediaPipe)
- Ideogram V2
- Kling Lip Sync Video with Text
- SamplerDPMAdaptative
- Ideogram V3
- Runway Image to Video (Gen4 Turbo)
- ByteDance Text to Video
- ChromaRadianceOptions
- ModelSamplingAuraFlow
- Wan 2.7 Reference to Video
- Load Image (from Folder)
- Convert Text to Lowercase (DEPRECATED)
- Kling Text to Video (Camera Control)
- Run MoGe Panorama Inference
- Concatenate Audio
- OpenAI DALL·E 3
- Combine Masks
- SamplerSEEDS2
- SamplerEulerCFG++
- Preview 3D (Advanced)
- Voxel to Mesh (Basic) (DEPRECATED)
- Apply Qwen Image DiffSynth ControlNet
- LTXVAddGuide
- PiD Conditioning
- Resolution Selector
- Edit Model Reference Method
- Flux.1 Fill Image
- Preview Mask
- Flux2Scheduler
- Save Audio (FLAC) (DEPRECATED)
- TripoSplat Conditioning
- Flux 1.1 [pro] Ultra Image
- Hunyuan3D: Model to UV
- Recraft Text to Image
- Load LoRA (Model and CLIP)
- SV3D_Conditioning
- ModelMergeFlux1
- Stitch Images
- Ideogram 4 Scheduler
- LTXV Preprocess
- Extract Text
- Rodin 3D Generate - Sketch Generate
- Rodin 3D Gen-2.5 - Image to 3D
- TripoSplat Preprocess Image
- Runway First-Last-Frame to Video
- Stability AI Audio Inpaint
- ElevenLabs Text to Sound Effects
- Save Audio (Advanced)
- Get IC-LoRA Parameters
- Cond Pair Combine
- Luma Text to Video
- Vidu Reference To Video Generation
- Cond Pair Set Default Combine
- Wan Image to Image
- Apply ControlNet
- Match Text
- Topaz Image Enhance
- CFGZeroStar
- OptimalStepsScheduler
- Empty Ace Step 1.0 Latent Audio
- Rebatch Images
- Epsilon Scaling
- Blend Images
- Scale Image to Max Dimension
- Add Text Prefix (DEPRECATED)
- StableCascade_StageC_VAEEncode
- ModelMergeSDXL
- Blur Image
- SkipLayerGuidanceDiTSimple
- Trim Video
- Set Reference Latent
- RescaleCFG
- ImageOnlyCheckpointSave
- Upscale Latent
- Timesteps Range
- Compare Images
- Save Image
- Load Depth Anything 3
- Batch Images (DEPRECATED)
- InpaintModelConditioning
- Replace Video Latent Frames
- SamplerDPMPP_2M_SDE
- Grok Image
- Resize And Pad Image
- Get Image Size
- Train LoRA
- BasicScheduler
- Crop Image
- Add Noise to Image
- UNetSelfAttentionMultiply
- ModelMergeCosmos7B
- EmptyHunyuanImageLatent
- Kling 2.6 Text to Video with Audio
- Preview Splat
- Vidu Text To Video Generation
- VOIDInpaintConditioning
- Convert Text Case
- Custom Combo
- LatentOperationSharpen
- Normalized Attention Guidance
- Create Camera Info
- Conditioning (Set Area with Percentage for Video)
- Kling 3.0 Image
- OpenAI ChatGPT
- Quiver Text to SVG
- PatchModelAddDownscale (Kohya Deep Shrink)
- Extract Mesh from Splat
- Upscale Image (using Model)
- LatentBatchSeedBehavior
- ModelMergeSimple
- TSR - Temporal Score Rescaling
- Load Latent Upscale Model
- WanMoveTracksFromCoords
- Render Splat
- Seed
- Preview Image
- KarrasScheduler
- Flux Erase Image
- VAE Encode Audio
- SDPose Face Bounding Boxes
- Load Hypernetwork
- Convert MoGe Point Map to Mesh
- Flux.1 Kontext [max] Image
- Kling Dual Character Video Effects
- Set Reference Audio
- ByteDance Reference Images to Video
- Load CLIP (Quadruple)
- Create List
- HiDream-O1 Reference Images
- LazyCache
- ScaleROPE
- Dual CFG Guider
- Save Animated PNG
- Empty HiDream-O1 Latent Image
- HiDream-O1 Patch Seam Smoothing
- Meshy: Multi-Image to Model
- TripoSplat Sampling Preview
- Magnific Image Upscale (Creative)
- ModelMergeAuraflow
- Hunyuan3D: Smart Topology
- Kling Image(First Frame) to Video
- Kling Video Extend
- Differential Diffusion
- Kling 2.6 Image(First Frame) to Video with Audio
- Luma Image to Image
- OpenAI ChatGPT Advanced Options
- PixVerse Image to Video
- Recraft Crisp Upscale Image
- Sonilo Video to Music
- Meshy: Animate Model
- MiniMax Image to Video
- Quiver Image to SVG
- WanTrackToVideo
- Stability AI Text To Audio
- Grok Video
- FreSca
- Vidu2 Reference-to-Video Generation
- Wan Image to Video
- FlashVSR Video Upscale
- Save Image-Text (to Folder)
- Run MoGe Inference
- Luma Ray 3.2 Video Reframe
- Meshy: Texture Model
- Runway Aleph2 Prompt Image
- Vidu Multi-Frame Video Generation
- Kling 3.0 Video
- Tripo P1: Text to Model
- Tripo P1: Image to Model
- Load Audio
- PolyexponentialScheduler
- Flip Image
- Load Checkpoint Image Only (img2vid model)
- Magnific Image Style Transfer
- Join Audio Channels
- Adjust Audio Volume
- WanVaceToVideo
- Wan22ImageToVideoLatent
- WanInfiniteTalkToVideo
- CFGNorm
- Google Gemini
- CLIP Text Encode (Hunyuan Image)
- Wan 2.7 Image to Video
- SDPose Keypoint Extractor
- Save Image (Advanced)
- Save WEBM
- Luma Reference
- WanSoundImageToVideoExtend
- WanHuMoImageToVideo
- KSamplerSelect
- TextEncodeAceStepAudio1.5
- CLIPAttentionMultiply
- Truncate Text
- SamplerLCMUpscale
- Color Picker
- Empty Latent Audio
- Strip Whitespace (DEPRECATED)
- CosmosPredict2ImageToVideoLatent
- Generate Text
- LaplaceScheduler
- Load Latent
- CLIP Text Encode (PixArt Alpha)
- Create Hook LoRA (MO)
- Flux Virtual Try-On
- Record Audio
- Combine Hooks [4]
- Set CLIP Hooks
- Detect Face Landmarks (MediaPipe)
- Runway Aleph2 Video to Video
- Hunyuan Latent Refiner
- Convert DA3 Geometry to Mesh
- Flux.1 Kontext [pro] Image
- Flux.2 [pro] Image
- Bria Remove Image Background
- Meshy: Text to Model
- Kling Avatar 2.0
- LTXV Image To Video
- SamplingPercentToSigma
- Dual Model CFG Guider
- Merge List of Tiles to Image
- Resolution Bucket
- Vidu Image To Video Generation
- Vidu Q3 Start/End Frame-to-Video Generation
- Wan Text to Image
- Luma UNI-1 Image Edit
- Flip Latent
- T5 Tokenizer Options
- AlignYourStepsScheduler
- Substring
- Draw BBoxes
- ConditioningSetTimestepRange
- Stability AI Upscale Conservative
- Vidu Q3 Text-to-Video Generation
- PixVerse Transition Video
- Rodin 3D Generate - Gen-2 Generate
- Recraft V4 Text to Image
- Recraft Style - Infinite Style Library
- Wan22FunControlToVideo
- WanSCAILToVideo
- Stability AI Upscale Creative
- Stability AI Upscale Fast
- Tripo: Texture model
- ByteDance Image
- ElevenLabs Speech to Speech
- Meshy: Rig Model
- Recraft Style - Realistic Image
- Reve Image Create
- Google Veo 2 Video Generation
- Vidu2 Image-to-Video Generation
- Bernini Conditioning
- VOIDWarpedNoise
- WanDancerPadKeyframesList
- ElevenLabs Speech to Text
- Recraft Controls
- WaveSpeed Image Upscale
- Save 3D Model
- Kling Lip Sync Video with Audio
- Kling 3.0 Omni Edit Video
- Runway Image to Video (Gen3a Turbo)
- Preview as Text
- KSampler (Advanced)
- Apply Style Model
- ModelSamplingSD3
- Recraft V4 Text to Vector
- ModelSamplingLTXV
- Create Hook Model as LoRA
- Grok Image Edit
- Kling Virtual Try On
- KSampler
- Adjust Contrast
- CLIPMergeAdd
- LatentSubtract
- Batch Masks
- Crop Image (Center)
- WanFunInpaintToVideo
- Create Video
- Select CLIP Device
- Trim Text
- CLIP Text Encode (SDXL)
- StableZero123_Conditioning_Batched
- NormalizeVideoLatentStart
- Load Diffusion Model
- Or
- WanDancerEncodeAudio
- Beeble SwitchX Video Edit
- Resize Image/Mask
- Transfer Color
- Shuffle Images List
- Add Text Suffix (DEPRECATED)
- Render MoGe Geometry
- Visualize Face Landmarks (MediaPipe)
- Magnific Image Upscale (Precise V2)
- WanSoundImageToVideo
- WanCameraEmbedding
- Plot Loss Graph
- SDPose Draw Keypoints
- Reve Image Remix
- Cond Set Props Combine
- MiniMax Hailuo Video
- Tripo: Text to Model
- Ideogram V1
- Vidu Q3 Image-to-Video Generation
- HappyHorse Reference to Video
- UNetCrossAttentionMultiply
- Load 3D (Advanced)
- Text String (DEPRECATED)
- Load Image (from Outputs)
- GITSScheduler
- TextEncodeHunyuanVideo_ImageToVideo
- Hunyuan Video 1.5 Super Resolution
- Create Hook Keyframes From Floats
- TorchCompileModel
- Load LTXV Audio VAE
- LTXVImgToVideo
- LTXVScheduler
- LTXVCropGuides
- Cond Set Props
- Empty Flux 2 Latent
- Cond Pair Set Props
- HitPaw Video Enhance
- Make Training Dataset
- Hunyuan Video 15 Latent Upscale With Model
- Load Audio Encoder
- Load Background Removal Model
- Create SCAIL-2 Colored Mask
- Cond Pair Set Props Combine
- Wan Context Windows
- Anthropic Claude
- FluxGuidance
- FluxKontextImageScale
- EmptyMochiLatentVideo
- Make Image Grid
- Tangential Damping CFG
- Nano Banana (Google Gemini Image)
- Bria Remove Video Background
- Grok Video Edit
- Luma Ray 3.2 Video Edit
- Luma Ray 3.2 Extend Video
- PixVerse Text to Video
- Load ControlNet Model (diff)
- Audio Equalizer (3-Band)
- LTXV Context Windows
- Kling Motion Control
- Vidu2 Text-to-Video Generation
- Load Image-Text (from Folder)
- SAM3 Detect
- SamplerDPMPP_3M_SDE
- WanPhantomSubjectToVideo
- ModelMergeSD3_2B
- Hunyuan3Dv2ConditioningMultiView
- HappyHorse Video Edit
- ByteDance Seedream 4.5 & 5.0
- AudioEncoderEncode
- LatentCut
- Flux.1 Expand Image
- Context Windows (Manual)
- Wan Text to Video
- Webcam Capture
- Preview Audio
- ModelMergeWAN2_1
- Split Audio Channels
- EmptyLatentHunyuan3Dv2
- Replace Text (Regex)
- Recraft Replace Background
- Tripo: Refine Draft model
- Tripo: Rig model
- Google Veo 3 First-Last-Frame to Video
- Vidu2 Start/End Frame-to-Video Generation
- Wan Reference to Video
- Wan 2.7 Video Continuation
- Wan 2.7 Video Edit
- Luma Concepts
- Split Image into List of Tiles
- Crop Image (Random)
- HunyuanImageToVideo
- Apply Controlnet with VAE
- WanCameraImageToVideo
- Transform Splat
- TripoSplat Decode
- HitPaw General Image Enhance
- Hunyuan3D: 3D Part
- Ideogram V4
- Recraft Image to Image
- Luma UNI-1 Image
- Meshy: Image to Model
- OpenAI DALL·E 2
- OpenAI Sora - Video (DEPRECATED)
- Set Union ControlNet Type
- Empty HunyuanVideo 1.0 Latent
- CLIPTextEncodeFlux
- Load LTXV Audio Text Encoder
- Wan 2.7 Text to Video
- Generate LTX2 Prompt
- Generate Video Tracks
- Vidu Video Extension
- VAESave
- VPScheduler
- SamplerER_SDE
- Input Text
- Int
- If/Else Switch
- Gemini Input Files
- Recraft Remove Background
- Runway Text to Image
- Topaz Video Enhance (Legacy)
- Tripo: Import Model
- WanImageToVideo
- WanMoveTrackToVideo
- ARVideoI2V
- SAM3 Track Preview
- Luma Ray 3.2 Keyframes to Video
- Recraft Image Inpainting
- Rodin 3D Generate - Smooth Generate
- Tripo: Multiview to Model
- Vidu Start End To Video Generation
- Save Image (to Folder) (DEPRECATED)
- ModelMergeCosmos14B
- WanFirstLastFrameToVideo
- CLIP Text Encode (Kandinsky 5)
- Create 3D File (from Splat)
- Empty Ace Step 1.5 Latent Audio
- Load LoRA (Bypass, Model Only) (for debugging)
- Remove Background
- Krea 2 Style Reference
- LTXV Text To Video
- Magnific Image Skin Enhancer
- Recraft Vectorize Image
- Bria Video Replace Background
- Save Video
- Grok Video Extend
- Rodin 3D Generate - Regular Generate
- Tripo: Image to Model
- Run Depth Anything 3
- EmptyChromaRadianceLatentImage
- Conditioning (Set Area with Percentage)
- Load Diffusers Model (DEPRECATED)
- LatentApplyOperationCFG
- Batch Latents
- Feather Mask
- Grow Mask
- Porter-Duff Image Composite
- CFG Override
- Runway Aleph2 Keyframe
- Stability AI Stable Image Ultra
- Curve Editor
- ByteDance Create Image Asset
- Load CLIP (Dual)
- StableZero123_Conditioning
- Load CLIP
- Sampler AR Video
- EmptyLTXVLatentVideo
- CLIP Text Encode (Lumina 2)
- Apply ControlNet Inpainting (AliMama)
- ByteDance Create Video Asset
- Reve Image Edit
- FreeU
- Load Training Dataset
- CosmosImageToVideoLatent
- Boolean
- Not
- Load MoGe Model
- Stability AI Stable Diffusion 3.5 Image
- Apply GLIGEN Text Box
- PerturbedAttentionGuidance
- Video Linear CFG Guidance
- Load PhotoMaker Model
- VAE Decode Audio
- ConditioningStableAudio
- CLIP Text Encode (SD3)
- ElevenLabs Voice Selector
- Google Gemini
- Pad Image for Outpainting
- SamplerLCM
- Empty Audio
- Float
- VAE Encode
- unCLIPConditioning
- Load Model Patch
- Rotate Latent
- ModelMergeSD2
- Select Model Device
- Crop Mask
- Bria Video Green Screen
- Tripo: Convert model
- Hunyuan3D: 3D Texture Edit
- CLIP Vision Encode
- ElevenLabs Text to Speech
- Conditioning (Average)
- ExtendIntermediateSigmas
- FluxDisableGuidance
- Flux.2 Image
- Hunyuan3Dv2Conditioning
- Create Hook Keyframes Interp.
- Kandinsky5ImageToVideo
- CLIP Text Encode (Prompt)
- ByteDance Seedance 2.0 Text to Video
- Save Training Dataset
- VOIDWarpedNoiseSource
- SkipLayerGuidanceDiT
- EmptyCosmosLatentVideo
- TextEncodeBooguEdit
- Google Veo 3 Video Generation
- Load CLIP Vision
- Conditioning Zero Out
- LatentApplyOperation
- Trim Audio Duration
- Kling 3.0 Omni Image
- Conditioning (Multiply)
- ModelSave
- Recraft Text to Vector
- WanFunControlToVideo
- Luma Ray 3.2 Text to Video
- Luma Ray 3.2 Keyframe
- Recraft Creative Upscale Image
- MiniMax Text to Video
- Conditioning (Combine)
- Basic Guider
- Kling Start-End Frame to Video
- Image Composite Masked
- Latent Composite Masked
- ElevenLabs Voice Isolation
- Kling Image to Video (Camera Control)
- Luma Text to Image
- Kling Camera Controls
- And
- Scale Image to Total Pixels
- Run SAM3 Video Track
- Apply Z-Image Fun ControlNet
- Crop By Bounding Boxes
- Load Checkpoint With Config (DEPRECATED)
- Convert Text to Uppercase (DEPRECATED)
- SAM3 Track to Mask
- Shuffle Pairs of Image-Text
- SVD_img2vid_Conditioning
README
ComfyUI
The most powerful and modular AI engine for content creation.
<!-- Workaround to display total user from https://github.com/badges/shields/issues/4500#issuecomment-2060079995 --> <img width="1590" height="795" alt="ComfyUI Screenshot" src="https://github.com/user-attachments/assets/36e065e0-bfae-4456-8c7f-8369d5ea48a2" /> <br> </div>ComfyUI is the AI creation engine for visual professionals who demand control over every model, every parameter, and every output. Its powerful and modular node graph interface empowers creatives to generate images, videos, 3D models, audio, and more...
- ComfyUI natively supports the latest open-source state of the art models.
- API nodes provide access to the best closed source models such as Nano Banana, Seedance, Hunyuan3D, etc.
- It is available on Windows, Linux, and macOS, locally with our desktop application, our portable install or on our cloud.
- The most sophisticated workflows can be exposed through a simple UI thanks to App Mode.
- It integrates seamlessly into production pipelines with our API endpoints.
Get Started
Local
Desktop Application
- The easiest way to get started.
- Available on Windows & macOS.
Windows Portable Package
- Get the latest commits and completely portable.
- Available on Windows.
Manual Install
Supports all operating systems and GPU types (NVIDIA, AMD, Intel, Apple Silicon, Ascend).
Cloud
Comfy Cloud
- Our official paid cloud version for those who can't afford local hardware.
Examples
See what ComfyUI can do with the newer template workflows or old example workflows.
Features
- Nodes/graph/flowchart interface to experiment and create complex Stable Diffusion workflows without needing to code anything.
- NOTE: There are many more models supported than the list below, if you want to see what is supported see our templates list inside ComfyUI.
- Image Models
- SD1.x, SD2.x (unCLIP)
- SDXL, SDXL Turbo
- Stable Cascade
- SD3 and SD3.5
- Pixart Alpha and Sigma
- AuraFlow
- HunyuanDiT
- Flux
- Lumina Image 2.0
- HiDream
- Qwen Image
- Hunyuan Image 2.1
- Flux 2
- Z Image
- Ernie Image
- Image Editing Models
- Video Models
- Audio Models
- 3D Models
- Asynchronous Queue system
- Many optimizations: Only re-executes the parts of the workflow that changes between executions.
- Smart memory management: can automatically run large models on GPUs with as low as 1GB vram with smart offloading.
- Works even if you don't have a GPU with:
--cpu(slow) - Can load ckpt and safetensors: All in one checkpoints or standalone diffusion models, VAEs and CLIP models.
- Safe loading of ckpt, pt, pth, etc.. files.
- Embeddings/Textual inversion
- Loras (regular, locon and loha)
- Hypernetworks
- Loading full workflows (with seeds) from generated PNG, WebP and FLAC files.
- Saving/Loading workflows as Json files.
- Nodes interface can be used to create complex workflows like one for Hires fix or much more advanced ones.
- Area Composition
- Inpainting with both regular and inpainting models.
- ControlNet and T2I-Adapter
- Upscale Models (ESRGAN, ESRGAN variants, SwinIR, Swin2SR, etc...)
- GLIGEN
- Model Merging
- LCM models and Loras
- Latent previews with TAESD
- Works fully offline: core will never download anything unless you want to.
- Optional API nodes to use paid models from external providers through the online Comfy API disable with:
--disable-api-nodes - Config file to set the search paths for models.
Workflow examples can be found on the Examples page
Release Process
ComfyUI follows a weekly release cycle targeting Monday but this regularly changes because of model releases or large changes to the codebase. There are three interconnected repositories:
-
- Releases a new major stable version (e.g., v0.7.0) roughly every 2 weeks.
- Starting from v0.4.0 patch versions will be used for fixes backported onto the current stable release.
- Minor versions will be used for releases off the master branch.
- Patch versions may still be used for releases on the master branch in cases where a backport would not make sense.
- Commits outside of the stable release tags may be very unstable and break many custom nodes.
- Serves as the foundation for the desktop release
-
- Builds a new release using the latest stable core version
-
- Every 2+ weeks frontend updates are merged into the core repository
- Features are frozen for the upcoming core release
- Development continues for the next release cycle
Shortcuts
| Keybind | Explanation |
|------------------------------------|--------------------------------------------------------------------------------------------------------------------|
| Ctrl + Enter | Queue up current graph for generation |
| Ctrl + Shift + Enter | Queue up current graph as first for generation |
| Ctrl + Alt + Enter | Cancel current generation |
| Ctrl + Z/Ctrl + Y | Undo/Redo |
| Ctrl + S | Save workflow |
| Ctrl + O | Load workflow |
| Ctrl + A | Select all nodes |
| Alt + C | Collapse/uncollapse selected nodes |
| Ctrl + M | Mute/unmute selected nodes |
| Ctrl + B | Bypass selected nodes (acts like the node was removed from the graph and the wires reconnected through) |
| Delete/Backspace | Delete selected nodes |
| Ctrl + Backspace | Delete the current graph |
| Space | Move the canvas around when held and moving the cursor |
| Ctrl/Shift + Click | Add clicked node to selection |
| Ctrl + C/Ctrl + V | Copy and paste selected nodes (without maintaining connections to outputs of unselected nodes) |
| Ctrl + C/Ctrl + Shift + V | Copy and paste selected nodes (maintaining connections from outputs of unselected nodes to inputs of pasted nodes) |
| Shift + Drag | Move multiple selected nodes at the same time |
| Ctrl + D | Load default graph |
| Alt + + | Canvas Zoom in |
| Alt + - | Canvas Zoom out |
| Ctrl + Shift + LMB + Vertical drag | Canvas Zoom in/out |
| P | Pin/Unpin selected nodes |
| Ctrl + G | Group selected nodes |
| Q | Toggle visibility of the queue |
| H | Toggle visibility of history |
| R | Refresh graph |
| F | Show/Hide menu |
| . | Fit view to selection (Whole graph when nothing is selected) |
| Double-Click LMB | Open node quick search palette |
| Shift + Drag | Move multiple wires at once |
| Ctrl + Alt + LMB | Disconnect all wires from clicked slot |
Ctrl can also be replaced with Cmd instead for macOS users
Installing
Windows Portable
There is a portable standalone build for Windows that should work for running on Nvidia GPUs or for running on your CPU only on the releases page.
Direct link to download
Simply download, extract with 7-Zip or with the windows explorer on recent windows versions and run. For smaller models you normally only need to put the checkpoints (the huge ckpt/safetensors files) in: ComfyUI\models\checkpoints but many of the larger models have multiple files. Make sure to follow the instructions to know which subfolder to put them in ComfyUI\models\
If you have trouble extracting it, right click the file -> properties -> unblock
The portable above currently comes with python 3.13 and pytorch cuda 13.0. Update your Nvidia drivers if it doesn't start.
All Official Portable Downloads:
Portable for Nvidia GPUs (supports 20 series and above).
Portable for Nvidia GPUs with pytorch cuda 12.6 and python 3.12 (Supports Nvidia 10 series and older GPUs).
How do I share models between another UI and ComfyUI?
See the Config file to set the search paths for models. In the standalone windows build you can find this file in the ComfyUI directory. Rename this file to extra_model_paths.yaml and edit it with your favorite text editor.
comfy-cli
You can install and start ComfyUI using comfy-cli:
pip install comfy-cli
comfy install
Manual Install (Windows, Linux)
Python 3.14 works but some custom nodes may have issues. The free threaded variant works but some dependencies will enable the GIL so it's not fully supported.
Python 3.13 is very well supported. If you have trouble with some custom node dependencies on 3.13 you can try 3.12
torch 2.5 is minimally supported but using a newer version is extremely recommended. Some features and optimizations might only work on newer versions. We generally recommend using the latest major version of pytorch with the latest cuda version unless it is less than 2 weeks old. If your pytorch is more than 6 months old, please update it.
Instructions:
Git clone this repo.
Put your SD checkpoints (the huge ckpt/safetensors files) in: models/checkpoints
Put your VAE in: models/vae
AMD GPUs (Linux)
AMD users can install rocm and pytorch with pip if you don't have it already installed, this is the command to install the stable version:
pip install torch torchvision torchaudio --index-url https://download.pytorch.org/whl/rocm7.2
This is the command to install the nightly with ROCm 7.2 which might have some performance improvements:
pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/rocm7.2
AMD GPUs (Experimental: Windows and Linux), RDNA 3, 3.5 and 4 only.
These have less hardware support than the builds above but they work on windows. You also need to install the pytorch version specific to your hardware.
RDNA 3 (RX 7000 series):
pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx110X-all/
RDNA 3.5 (Strix halo/Ryzen AI Max+ 365):
pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx1151/
RDNA 4 (RX 9000 series):
pip install --pre torch torchvision torchaudio --index-url https://rocm.nightlies.amd.com/v2/gfx120X-all/
Intel GPUs (Windows and Linux)
Intel Arc GPU users can install native PyTorch with torch.xpu support using pip. More information can be found here
- To install PyTorch xpu, use the following command:
pip install torch torchvision torchaudio --index-url https://download.pytorch.org/whl/xpu
This is the command to install the Pytorch xpu nightly which might have some performance improvements:
pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/xpu
NVIDIA
Nvidia users should install stable pytorch using this command:
pip install torch torchvision torchaudio --extra-index-url https://download.pytorch.org/whl/cu130
This is the command to install pytorch nightly instead which might have performance improvements.
pip install --pre torch torchvision torchaudio --index-url https://download.pytorch.org/whl/nightly/cu132
Troubleshooting
If you get the "Torch not compiled with CUDA enabled" error, uninstall torch with:
pip uninstall torch
And install it again with the command above.
Dependencies
Install the dependencies by opening your terminal inside the ComfyUI folder and:
pip install -r requirements.txt
After this you should have everything installed and can proceed to running ComfyUI.
Others:
Apple Mac silicon
You can install ComfyUI in Apple Mac silicon (M1, M2, M3 or M4) with any recent macOS version.
- Install pytorch nightly. For instructions, read the Accelerated PyTorch training on Mac Apple Developer guide (make sure to install the latest pytorch nightly).
- Follow the ComfyUI manual installation instructions for Windows and Linux.
- Install the ComfyUI dependencies. If you have another Stable Diffusion UI you might be able to reuse the dependencies.
- Launch ComfyUI by running
python main.py
Note: Remember to add your models, VAE, LoRAs etc. to the corresponding Comfy folders, as discussed in ComfyUI manual installation.
Ascend NPUs
For models compatible with Ascend Extension for PyTorch (torch_npu). To get started, ensure your environment meets the prerequisites outlined on the installation page. Here's a step-by-step guide tailored to your platform and installation method:
- Begin by installing the recommended or newer kernel version for Linux as specified in the Installation page of torch-npu, if necessary.
- Proceed with the installation of Ascend Basekit, which includes the driver, firmware, and CANN, following the instructions provided for your specific platform.
- Next, install the necessary packages for torch-npu by adhering to the platform-specific instructions on the Installation page.
- Finally, adhere to the ComfyUI manual installation guide for Linux. Once all components are installed, you can run ComfyUI as described earlier.
Cambricon MLUs
For models compatible with Cambricon Extension for PyTorch (torch_mlu). Here's a step-by-step guide tailored to your platform and installation method:
- Install the Cambricon CNToolkit by adhering to the platform-specific instructions on the Installation
- Next, install the PyTorch(torch_mlu) following the instructions on the Installation
- Launch ComfyUI by running
python main.py
Iluvatar Corex
For models compatible with Iluvatar Extension for PyTorch. Here's a step-by-step guide tailored to your platform and installation method:
- Install the Iluvatar Corex Toolkit by adhering to the platform-specific instructions on the Installation
- Launch ComfyUI by running
python main.py
ComfyUI-Manager
ComfyUI-Manager is an extension that allows you to easily install, update, and manage custom nodes for ComfyUI.
Setup
-
Install the manager dependencies:
pip install -r manager_requirements.txt -
Enable the manager with the
--enable-managerflag when running ComfyUI:python main.py --enable-manager
Command Line Options
| Flag | Description |
|------|-------------|
| --enable-manager | Enable ComfyUI-Manager |
| --enable-manager-legacy-ui | Use the legacy manager UI instead of the new UI (implies --enable-manager) |
| --disable-manager-ui | Disable the manager UI and endpoints while keeping background features like security checks and scheduled installation completion (requires --enable-manager) |
Running
python main.py
For AMD cards not officially supported by ROCm
Try running it with this command if you have issues:
For 6700, 6600 and maybe other RDNA2 or older: HSA_OVERRIDE_GFX_VERSION=10.3.0 python main.py
For AMD 7600 and maybe other RDNA3 cards: HSA_OVERRIDE_GFX_VERSION=11.0.0 python main.py
AMD ROCm Tips
You can try setting this env variable PYTORCH_TUNABLEOP_ENABLED=1 which might speed things up at the cost of a very slow initial run.
Notes
Only parts of the graph that have an output with all the correct inputs will be executed.
Only parts of the graph that change from each execution to the next will be executed, if you submit the same graph twice only the first will be executed. If you change the last part of the graph only the part you changed and the part that depends on it will be executed.
Dragging a generated png on the webpage or loading one will give you the full workflow including seeds that were used to create it.
You can use () to change emphasis of a word or phrase like: (good code:1.2) or (bad code:0.8). The default emphasis for () is 1.1. To use () characters in your actual prompt escape them like \( or \).
You can use {day|night}, for wildcard/dynamic prompts. With this syntax "{wild|card|test}" will be randomly replaced by either "wild", "card" or "test" by the frontend every time you queue the prompt. To use {} characters in your actual prompt escape them like: \{ or \}.
Dynamic prompts also support C-style comments, like // comment or /* comment */.
To use a textual inversion concepts/embeddings in a text prompt put them in the models/embeddings directory and use them in the CLIPTextEncode node like this (you can omit the .pt extension):
embedding:embedding_filename.pt
How to show high-quality previews?
Use --preview-method auto to enable previews.
The default installation includes a fast latent preview method that's low-resolution. To enable higher-quality previews with TAESD, download the taesd_decoder.pth, taesdxl_decoder.pth, taesd3_decoder.pth and taef1_decoder.pth and place them in the models/vae_approx folder. Once they're installed, restart ComfyUI and launch it with --preview-method taesd to enable high-quality previews.
How to use TLS/SSL?
Generate a self-signed certificate (not appropriate for shared/production use) and key by running the command: openssl req -x509 -newkey rsa:4096 -keyout key.pem -out cert.pem -sha256 -days 3650 -nodes -subj "/C=XX/ST=StateName/L=CityName/O=CompanyName/OU=CompanySectionName/CN=CommonNameOrHostname"
Use --tls-keyfile key.pem --tls-certfile cert.pem to enable TLS/SSL, the app will now be accessible with https://... instead of http://....
Note: Windows users can use alexisrolland/docker-openssl or one of the 3rd party binary distributions to run the command example above. <br/><br/>If you use a container, note that the volume mount
-vcan be a relative path so... -v ".\:/openssl-certs" ...would create the key & cert files in the current directory of your command prompt or powershell terminal.
Support and dev channel
Discord: Try the #help or #feedback channels.
Matrix space: #comfyui_space:matrix.org (it's like discord but open source).
See also: https://www.comfy.org/
psst — we're hiring! Help build ComfyUI: comfy.org/careers
Frontend Development
As of August 15, 2024, we have transitioned to a new frontend, which is now hosted in a separate repository: ComfyUI Frontend. The compiled JS files (from TS/Vue) are published to pypi and installed as a dependency in ComfyUI.
Reporting Issues and Requesting Features
For any bugs, issues, or feature requests related to the frontend, please use the ComfyUI Frontend repository. This will help us manage and address frontend-specific concerns more efficiently.
Using the Latest Frontend
The new frontend is now the default for ComfyUI. However, please note:
- The frontend in the main ComfyUI repository is updated fortnightly.
- Daily releases are available in the separate frontend repository.
To use the most up-to-date frontend version:
-
For the latest daily release, launch ComfyUI with this command line argument:
--front-end-version Comfy-Org/ComfyUI_frontend@latest -
For a specific version, replace
latestwith the desired version number:--front-end-version Comfy-Org/ComfyUI_frontend@1.2.2
This approach allows you to easily switch between the stable fortnightly release and the cutting-edge daily updates, or even specific versions for testing purposes.
QA
Which GPU should I buy for this?
ComfyUI is ready to run
It's one of 95 extensions already installed on ComfyICU — nothing to clone, nothing to reconcile. Bring a workflow and you're billed for GPU seconds, not idle time.