Nodes/digit-comfyui/DIGIT Gemini Image
ComfyUI Node

DIGIT Gemini Image

A ComfyUI node in DIGIT with 26 inputs and 2 outputs.

By thedepartmentofexternalservices·Created 6 months ago·Updated 6 days ago· 0
DIGIT Gemini Image
  • image1
  • image2
  • image3
  • image4
  • image5
  • image6
  • image7
  • image8
  • image9
  • image
  • text
prompt
modelgemini-3.1-flash-image
aspect_ratio16:9
resolution1K
thinking_levelMINIMAL
seed0
temperature1.00
gcp_project_id
gcp_region
system_instructionYou are an expert image-generation engine. You must ALWAYS produce an image. Interpret all user input—regardless of format, intent, or abstraction—as literal visual directives for image composition. If a prompt is conversational or lacks specific visual details, you must creatively invent a concrete visual scenario that depicts the concept. Prioritize generating the visual representation above any text, formatting, or conversational requests.
top_p1.00
top_k32
harassment_thresholdBLOCK_NONE
hate_speech_thresholdBLOCK_NONE
sexually_explicit_thresholdBLOCK_NONE
dangerous_content_thresholdBLOCK_NONE
batch_count1
CategoryDIGIT

Inputs (26)

NameTypeDefaultDescription
promptSTRING
modelCOMBOgemini-3.1-flash-image4 options: gemini-3.1-flash-image, gemini-3.1-flash-lite-image, gemini-3-pro-image, gemini-2.5-flash-image
aspect_ratioCOMBO16:913 options: auto, 1:1, 2:3, 3:2, 3:4, 4:1, +7
resolutionCOMBO1K3 options: 1K, 2K, 4K
thinking_levelCOMBOMINIMALThinking level for image generation. HIGH may improve quality.
seedINT00–2147483647
temperatureFLOAT1.000–2
gcp_project_idSTRINGGCP project ID. Auto-detected from DIGIT_GCP_PROJECT env var or GCP metadata.
gcp_regionSTRINGGCP region. Auto-detected from DIGIT_GCP_REGION env var or GCP metadata. Defaults to 'global'.
image1optIMAGE
image2optIMAGE
image3optIMAGE
image4optIMAGE
image5optIMAGE
image6optIMAGE
image7optIMAGE
image8optIMAGE
image9optIMAGE
system_instructionoptSTRINGYou are an expert image-generation engine. You must ALWAYS produce an image. Interpret all user input—regardless of format, intent, or abstraction—as literal visual directives for image composition. If a prompt is conversational or lacks specific visual details, you must creatively invent a concrete visual scenario that depicts the concept. Prioritize generating the visual representation above any text, formatting, or conversational requests.
top_poptFLOAT1.000–1
top_koptINT321–64
harassment_thresholdoptCOMBOBLOCK_NONE4 options: BLOCK_NONE, BLOCK_ONLY_HIGH, BLOCK_MEDIUM_AND_ABOVE, BLOCK_LOW_AND_ABOVE
hate_speech_thresholdoptCOMBOBLOCK_NONE4 options: BLOCK_NONE, BLOCK_ONLY_HIGH, BLOCK_MEDIUM_AND_ABOVE, BLOCK_LOW_AND_ABOVE
sexually_explicit_thresholdoptCOMBOBLOCK_NONE4 options: BLOCK_NONE, BLOCK_ONLY_HIGH, BLOCK_MEDIUM_AND_ABOVE, BLOCK_LOW_AND_ABOVE
dangerous_content_thresholdoptCOMBOBLOCK_NONE4 options: BLOCK_NONE, BLOCK_ONLY_HIGH, BLOCK_MEDIUM_AND_ABOVE, BLOCK_LOW_AND_ABOVE
batch_countoptINT11–128Number of images to generate. Each is a separate API call fired in parallel; results return as one IMAGE batch.

Outputs (2)

NameTypeDescription
imageIMAGE
textSTRING