ComfyUI Node: π¨ NVIDIA Captioner
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.
Category
NVIDIA/Vision
Inputs
image_directory STRING
api_key STRING
model
- meta/llama-3.2-11b-vision-instruct
- meta/llama-3.2-90b-vision-instruct
- meta/llama-3.2-11b-vision
- meta/llama-3.2-90b-vision
system_prompt_preset
- default
- women_influencer
- cartoon_character
- interior_design
- product_shot
custom_system_prompt STRING
prompt STRING
use_cache BOOLEAN
skip_existing_txt BOOLEAN
max_tokens INT
temperature FLOAT
top_p FLOAT
frequency_penalty FLOAT
presence_penalty FLOAT
max_retries INT
retry_delay FLOAT
Outputs
STRING
STRING
Extension: ComfyUI-NvidiaCaptioner
A ComfyUI node for generating rich, detailed captions for images using NVIDIA's vision models. Supports batch processing, multiple captioning styles, and includes built-in caching for efficient workflows.
Authored by theshubzworld
Looking for a different node?
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.