Nodes/ComfyUI-API-Toolkit/Gemini Vision Analysis
ComfyUI Node

Gemini Vision Analysis

A ComfyUI node in API Toolkit/Gemini/Image with 10 inputs and 1 output.

By IxMxAMAR·Created 5 months ago·Updated about a month ago· 1
Gemini Vision Analysis
  • ref_image_1
  • ref_image_2
  • ref_image_3
  • ref_image_4
  • text
api_key
modelgemini-3.1-flash-lite-preview
custom_model
promptDescribe this image.
system_instruction
temperature0.10
CategoryAPI Toolkit/Gemini/Image

Inputs (10)

NameTypeDefaultDescription
api_keySTRINGGemini API key. Leave blank to use GEMINI_API_KEY env var.
modelCOMBOgemini-3.1-flash-lite-previewGemini model for vision analysis.
custom_modelSTRINGOverride with a custom model ID.
promptSTRINGDescribe this image.What to ask about the image(s).
system_instructionoptSTRINGSystem instruction to guide analysis.
temperatureoptFLOAT0.100–2Controls randomness. Lower = more focused analysis.
ref_image_1optIMAGEFirst image to analyze.
ref_image_2optIMAGESecond image to analyze.
ref_image_3optIMAGEThird image to analyze.
ref_image_4optIMAGEFourth image to analyze.

Outputs (1)

NameTypeDescription
textSTRING