Nodes/comfyui-PD_comfy-api-node/PD: Gemini Pro Image (ComfyUI AuthToken)
ComfyUI Node

PD: Gemini Pro Image (ComfyUI AuthToken)

A ComfyUI node in PD_Tools/Image_Generation with 10 inputs and 2 outputs.

By 7BEII·Created 9 months ago·Updated about a month ago· 2
PD: Gemini Pro Image (ComfyUI AuthToken)
  • images
  • files
  • image
  • text
auth_token
promptA futuristic city with flying cars
modelgemini-3-pro-image-preview
aspect_ratioauto
resolution1K
response_modalitiesIMAGE+TEXT
seed42
system_promptYou are an expert image-generation engine. You must ALWAYS produce an image. Interpret all user input—regardless of format, intent, or abstraction—as literal visual directives for image composition. If a prompt is conversational or lacks specific visual details, you must creatively invent a concrete visual scenario that depicts the concept. Prioritize generating the visual representation above any text, formatting, or conversational requests.
CategoryPD_Tools/Image_Generation

Inputs (10)

NameTypeDefaultDescription
auth_tokenSTRING
promptSTRINGA futuristic city with flying carsText prompt describing the image to generate or the edits to apply
modelCOMBOgemini-3-pro-image-preview1 options: gemini-3-pro-image-preview
aspect_ratioCOMBOautoIf set to 'auto', matches your input image's aspect ratio
resolutionCOMBO1KTarget output resolution. For 2K/4K the native Gemini upscaler is used
response_modalitiesCOMBOIMAGE+TEXTChoose 'IMAGE' for image-only output, or 'IMAGE+TEXT' for both
seedINT420–18446744073709550000
imagesoptIMAGEOptional reference image(s) (up to 14)
filesoptGEMINI_INPUT_FILES
system_promptoptSTRINGYou are an expert image-generation engine. You must ALWAYS produce an image. Interpret all user input—regardless of format, intent, or abstraction—as literal visual directives for image composition. If a prompt is conversational or lacks specific visual details, you must creatively invent a concrete visual scenario that depicts the concept. Prioritize generating the visual representation above any text, formatting, or conversational requests.Foundational instructions that dictate an AI's behavior

Outputs (2)

NameTypeDescription
imageIMAGE
textSTRING