civitai-comfy-nodes
ComfyUI custom node pack exposing Civitai Orchestration recipe endpoints as generated nodes
Nodes (170)
Full songs from a prompt — ACE-Step music generation, no GPU needed
Turn audio into text for training sets, without captioning it by hand
The one node this whole pack runs on — your Civitai key, minus the pain
OpenAI-style chat completions with JSON, tools, and even image output — wired into your graph
A chat LLM inside your graph without writing a line of JSON
Pick a preprocessor, wire an image, chain your way to structure
Textual-inversion embeddings for the cloud — no trigger word required
One image in, transparent PNG out
Anima in the cloud — the anime model with actual prompt comprehension
Anima variants from an existing image — same character, new pose, cloud-rendered
Boogu-Image on the cloud — the surprise 2026 base that went head-to-head with Klein
Boogu-Image as an editor — a sentence to change what's in the frame
Boogu Turbo — the 4-step draft machine, cloud-served
ERNIE-Image on the cloud — the layout-and-text specialist, minus the hype cycle
ERNIE Turbo — the fast 8-step variant, as a cloud node
Pick a checkpoint, type a prompt, get an image
FLUX.1 variants from an existing image — img2img without the local model
Flux 2 Dev — the model nobody could run locally, now just a node
Flux 2 Dev as an editor — change what's in a photo with a sentence
HiDream-O1 — the no-VAE, 8B pixel-space model, rendered on someone else's GPU
HiDream-O1 without the 20GB download — and the 2048 default that tells you what you're in for
Instruction editing, no mask required
The full 8B editor, 40 steps and all — for when dev isn't enough
Civitai's comfy/base node runs the API checkpoint, not the aligned open weights
Eight steps and done — the cheap draft door into the 12B model
Civitai's mystery image model, 20 steps at cfg 4
Four steps at cfg 1 — the fastest draft this pack sells
The SD1.5 cloud node nobody needs — and the three reasons you might want it anyway
A variant node with a denoise knob and a required image
The 1024px node that makes the 'no GPU' case for the pack
CreateVariant at 1024px, image in and variant out
Fal's API shape, with style references and a creativity slider
The thinnest node in the pack, and honestly that's fine
The whole API in two inputs — image in, sentence, image out
Qwen2 / createImage, prompt expansion on by default
The node where the mask died, billed by the edit
The watermark-killer and character-consistency tool, minus the license scare
The strongest image model you'll never have to fit in VRAM
The node with all the knobs, for when you want Flux 2 without the 24GB install
The one checkpoint that generates and edits, now with an image input
The Civitai editImage node
Flux 2 flex text-to-image with no local GPU (and the dials to prove it)
Image-to-image editing without a local checkpoint
The cloud node that feels local
Hosted Flux 2 Klein editing with LoRAs and a negative prompt
Gemini 2.5 Flash text-to-image, a.k.a. the base Nano Banana, as a Comfy node
Give it a photo, get back a rewrite
The create-only option with a negative prompt
The Google node that checks facts before drawing
4K Gemini image generation, hosted
The pack's least-restrictive create node
Hosted image editing from xAI's autoregressive model
The museum piece of the Civitai pack
The pack's only node with a real mask input
The last-gen classic with style and quality knobs
OpenAI's current text-to-image, minus the key hassle
The input_fidelity dial nobody else has
The model that went viral, hosted
The hidden mask input on OpenAI's first-gen image editor
OpenAI's newest, and the one with real pixel controls
Edit with gpt-image-2 from your ComfyUI graph — the model you'll never download
The natural-language anime model, no 2B download
Re-render an image on Anima — plain-language restyle on the cloud
Flux.1 without the 24GB download — bring your own AIR stack
Re-render with the four-piece stack
Edit a photo with Flux.1's full stack, without a GPU that can hold it
Flux 2 Dev, the 50GB model you can finally run — because it runs on their machine
Flux 2 Dev img2img without the 50GB download
Instruction-edit a photo with Flux 2 Dev, on Civitai's fleet
Pick your 4B or 9B, skip the download
Restyle a picture with a 4B or 9B Flux 2
The cheap way to edit with Flux 2
Qwen-Image, the 20B text-rendering champ, on Civitai's fleet
Variant of a Qwen render, text included, on the cloud
Edit a picture with Qwen-Image's instruction following — and keep the text legible
10,000 checkpoints, none of them on your disk
Restyle a picture on any SD 1.5 checkpoint — no downloads, no VRAM
Your favorite checkpoint, rendered on Civitai's fleet
Re-render any image on an SDXL checkpoint you don't even have locally
Z-Image Base on the cloud — the 'SDXL 2.0' model, no 6B download
Paying Buzz to run the free model — Z-Image Turbo on Civitai's farm
Seedream, the closed model you can't run locally — as a Comfy node
Wan 2.2's small model — the one that fits on 8GB, in the cloud
Wan 2.2 — the last open Wan, on someone else's GPU
Wan 2.5 image-to-image — closed weights, open-ended editing
Wan 2.5 — the first Wan you literally cannot run locally
Wan 2.7 — the closed flagship, and the Buzz meter for the 'pro' tier
Wan 2.7 edit — where the closed flagship actually earns its keep
Train a Flux 2 Dev LoRA without owning an RTX 6000
Training a LoRA that edits — Flux 2 Dev edit, minus the 48GB card
Flux dev fast — the trainer with the training wheels left on
The entire kohya training UI, moved to the cloud
Musubi, the Wan trainer, without the Dockerfile
The cloud upscaler for people with no GPU to upscale on
Apply one LoRA straight from a Civitai URL
Auto-captions without installing a VLM — Civitai's captioner node
Civitai's own moderation brain, exposed as a node
The one node that makes the rest of the Civitai pack make sense
An image becomes a 3D mesh — Meshy, without installing anything 3D
Prompt to a 3D mesh, the way the demos make it look
An LLM that writes prompts the way each model actually wants them
Score your generated images against their prompts
One text box to speech — with optional voice cloning, no local model
Your reference audio becomes the voice
Qwen3 TTS with zero setup
Describe a voice in words, get it speaking — no samples, no speaker list
Train an ACE-Step music LoRA on Civitai's cloud — no local GPU required
Pick base or SFT, let the cloud sweat
Train an Anima LoRA on the cloud — the anime model that made people switch
Cloud LoRA training for Chroma and friends
Flux.1 LoRA training on the cloud — dev or schnell, you pick
4B or 9B, plus an edit-LoRA toggle
Qwen-Image LoRA training on the cloud, with a version picker
SDXL and SD 1.5 LoRA training on the cloud — the familiar dials, no GPU
Turn any audio into text (and word-level timestamps) without a local model
Upscale and interpolate a video in one cloud call — no GPU babysitting
Edit an existing clip with xAI Grok — prompt-driven video editing in the cloud
Make a still image move with xAI Grok — image-to-video on the cloud
Prompt to video with xAI's Grok — no local video model needed
Haiper with every dial exposed
Turn an image into motion with a dancing name
Point at a look, get a clip — no GPU required
Prompt in, clip out, your GPU stays asleep
Reimagine a clip with a prompt and reference images
Animate a single frame on Civitai's cloud
The extra aspect ratios are the story
The newer engine, all nine aspect ratios
Run Tencent's model on the fleet, pick your own checkpoint
The closed model you can't download, called from a node
Five video operations in one node, plus audio
LTX's company, hosted, no 22GB encoder download
Make the sound drive the picture, on Civitai's cloud
The 22B model without the VRAM grind
Canny-guided video editing on Civitai's fleet
Keep the story going past the clip, on Civitai's cloud
Pin the beginning and end, let the cloud animate between
Gentle restyle of footage, structure intact, on Civitai
The 19B synced-audio model, without the Gemma headache
Canny-guided edits on the 19B, hosted by Civitai
Continue a clip with the 19B, hosted by Civitai
Script a scene by pinning its two bookend frames
Civitai's cloud recipe node
Prompt in, clip out
ByteDance's video engine on Civitai's fleet
Sora image-to-video from your ComfyUI canvas — via Civitai's cloud
Prompt in, cloud-rendered clip out
Veo 3 inside ComfyUI — native audio and all, on Civitai's cloud
Start-frame to end-frame video on Civitai's cloud
Resolution and speed dials on Civitai's cloud
Wan 2.1 on Civitai's own fleet — width, height, LoRAs and all
Aspect ratio, steps, and prompt expansion
The 8GB-friendly model, cloud-hosted
720p motion without a local GPU
Full sampler control on Civitai's cloud
The quality-first open model, cloud-hosted
Open-model quality without the hardware
The first closed Wan, via fal on Civitai's cloud
1080p without the open weights
Multi-shot and background music
Keep a subject consistent across clips
The closed release with multi-shot
Wan 2.7 video editing without a GPU (you just pay for it)
Turn a still into motion with Wan 2.7, GPU not included
Wan 2.7 reference-to-video, cloud-hosted
The API-only Wan 2.7, wrapped as a ComfyUI node
Frame interpolation on Civitai's fleet — the smooth-video button
The Civitai fleet upscaler
Auto-tag your images the lazy way — WD tagging, hosted by Civitai
Ask Civitai's moderation bot to grade your prompt before you hit generate
'blocked or not?'
civitai-comfy-nodes
🔹 Now in early preview — installable from the Comfy Registry. Still under active development: nodes, APIs, and behavior may change without notice. Use at your own risk.
ComfyUI custom nodes for the Civitai Orchestration API. Run Civitai's cloud recipes — image/video/audio generation, upscaling, training, captioning, moderation — as nodes inside any local ComfyUI graph. No local GPU or model downloads needed; jobs run on Civitai's fleet and are billed in Buzz.
Install
From the Comfy Registry (recommended)
The pack is published to the Comfy Registry, so ComfyUI Manager can install and update it for you:
-
In ComfyUI Manager: open Manager → Custom Nodes Manager, search for Civitai Comfy Nodes (publisher
civitai), and click Install, then restart ComfyUI. -
With comfy-cli:
comfy node registry-install civitai-comfy-nodes
See the Comfy Registry docs for details.
From source
Clone (or unzip) into your ComfyUI custom_nodes directory:
cd ComfyUI/custom_nodes
git clone https://github.com/civitai/civitai-comfy-nodes.git
pip install -r civitai-comfy-nodes/requirements.txt # just `requests`
Authentication
Nodes resolve credentials in this order:
- A connected Civitai Auth node (explicit token / base URL / mature-content / timeout overrides)
- The
CIVITAI_API_TOKENenvironment variable (create an API key) - A stored API key (
~/.civitai/comfy-api-key) or OAuth login (~/.civitai/comfy-oauth.json, auto-refreshed) — both can be set from the Civitai sidebar's connect panel (no env var needed) - An interactive browser login (OAuth + PKCE) — requires
CIVITAI_OAUTH_CLIENT_IDto be configured
Headless/remote ComfyUI installs should use the env var.
Nodes
~160 nodes under the Civitai category, generated from the orchestration OpenAPI spec. The menu
is organized ecosystem-first — Civitai/<media>/<ecosystem>[/<engine>]/… — with the engine
(sdcpp/comfy) shown as a sub-level only when an ecosystem is reachable through more than one engine
(e.g. Civitai/Image/zImage › zImage / turbo / createImage, Civitai/Image/anima/sdcpp). Each
discriminator variant is its own node so it shows only the inputs that variant actually uses:
- Civitai/Image — Image Gen (one node per engine: Flux2, OpenAI, Google, Seedream, …), Upscaler, Background Removal
- Civitai/Video — Video Gen (one node per engine: Wan, Kling, Vidu, Veo3, LTX, Sora, …), Upscaler, Interpolation, Enhancement
- Civitai/Audio — Text To Speech, Transcription, Audio Captioning, ACE Step Audio
- Civitai/Text — Chat Completion (plus a simple single-turn wrapper), Prompt Enhancement, Media Captioning
- Civitai/Analysis — Media Rating, WD Tagging, XGuard Moderation
- Civitai/Training — Training, Image Resource Training
- Civitai/Misc — Poly Gen (3D mesh generation)
- Civitai/Loaders — Model Selector, LoRA Selector, Embedding Selector, ControlNet (see below)
Every node returns its media outputs as native Comfy types (IMAGE/VIDEO/AUDIO) plus
workflow_id and raw_json for debugging and cost inspection. Models and LoRAs are
referenced by AIR URNs (e.g.
urn:air:sdxl:checkpoint:civitai:101055@128078).
Models, LoRAs, ControlNets & embeddings
Recipe nodes expose their model references as typed sockets, not text widgets — model/vae
are CIVITAI_AIR, loras is CIVITAI_LORAS, embeddings is CIVITAI_EMBEDDINGS. You fill them by
wiring a Civitai/Loaders selector node; each selector has a 🔍 Browse Civitai button (a
searchable card grid of generation-capable models via a same-origin proxy to
civitai.com/api/v1/models) and shows an on-node preview (thumbnail + name) of its current
resource. Recipe nodes themselves have no Browse button — change a model by wiring a selector.
-
Civitai Model Selector — pick a model once; it serves two purposes through its two outputs:
- Choose the model a Civitai recipe node runs on. Wire its
airoutput into a recipe node'smodel/vaesocket. The job runs on Civitai's fleet, so nothing is downloaded locally. - Auto-download a model for a local loader. Wire its
pathoutput into any standard loader's file widget — e.g. drop it in front of a Load LoRA (lora_name) or Load Checkpoint (ckpt_name) node and it downloads the model into the matching ComfyUI folder for you, no manual file management. No need to replace the loader — just feed it.
The download happens only when
pathis connected, so use (1) stays cloud-only and never pulls files down. - Choose the model a Civitai recipe node runs on. Wire its
-
Civitai LoRA Selector — holds multiple LoRAs in one node: each row has an enable toggle, the model (pick/replace via Browse Civitai), a strength, and a keywords field; + Add LoRA appends another. Wire its
lorasoutput into a recipe node'sloras/additional_networksinput (chain another selector via thelorasinput to combine stacks). Wire MODEL + CLIP to also download & apply the enabled LoRAs locally. The keywords field is auto-filled from the LoRA's trained words purely as a reminder — a LoRA applies by its AIR + strength, so the generator ignores these words; to actually invoke the trained concept, paste them into your recipe node's prompt. -
Civitai Embedding Selector — pick textual-inversion embeddings; chain (
embeddings→embeddings) and wire into a recipe node'sembeddingsinput. Embeddings apply automatically (the sdcpp pipeline prepends each one to the positive prompt), so no trigger word is needed. To use a negative embedding, reference it by its model name in the recipe node's negative prompt — that places it on the negative side instead. (Only sdcpp ecosystems — sd1/sdxl — expose anembeddingsinput.) -
Civitai ControlNet — pick a preprocessor, weight, step range, optional control image; chain and wire into a recipe node's
control_netsinput.
chat messages and other freeform structures remain JSON text inputs.
Browse your generations
The Civitai sidebar tab (the logo icon in the left rail) lists your Civitai generation history —
every workflow you've run, across all media types and sources (web / API / ComfyUI) — pulled from
the orchestrator and scoped to your account. Filter by media kind, scroll to paginate, and click any
result for a lightbox. Pull a result back into the graph three ways: add to canvas (creates the
matching loader node — image→LoadImage, video/audio/3D→their loaders — wired to the imported file),
fill a selected loader node, or drag a thumbnail onto the canvas. If no credentials are
configured, the tab shows a connect panel (OAuth sign-in or paste an API key).
Development
Nodes are generated — never edit civitai_comfy_nodes/generated/ by hand. To change
node shapes, edit codegen/overrides.json (or the codegen itself) and regenerate:
python -m codegen.generate # regenerate from spec/v2-consumers.json
pytest tests -q # unit tests (no ComfyUI or network needed)
CIVITAI_API_TOKEN=... pytest -m e2e -o addopts="" tests/test_e2e.py # prod smoke test
To pick up orchestration API changes, see scripts/sync-spec.sh.