ComfyUI Node: 🟨 NVIDIA Captioner

Authored by theshubzworld

Created

Updated

1 stars

Run ComfyUI workflows without the setup

No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.

Category

NVIDIA/Vision

Inputs

image_directory STRING
api_key STRING
model
  • meta/llama-3.2-11b-vision-instruct
  • meta/llama-3.2-90b-vision-instruct
  • meta/llama-3.2-11b-vision
  • meta/llama-3.2-90b-vision
system_prompt_preset
  • default
  • women_influencer
  • cartoon_character
  • interior_design
  • product_shot
custom_system_prompt STRING
prompt STRING
use_cache BOOLEAN
skip_existing_txt BOOLEAN
max_tokens INT
temperature FLOAT
top_p FLOAT
frequency_penalty FLOAT
presence_penalty FLOAT
max_retries INT
retry_delay FLOAT

Outputs

STRING

STRING

Extension: ComfyUI-NvidiaCaptioner

A ComfyUI node for generating rich, detailed captions for images using NVIDIA's vision models. Supports batch processing, multiple captioning styles, and includes built-in caching for efficient workflows.

Authored by theshubzworld

Looking for a different node?

Run ComfyUI workflows without the setup

No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.

Learn more