Nodes/Qwen2.5-VL GGUF Nodes/🖼️ Local Image Analysis (GGUF)
ComfyUI Node

🖼️ Local Image Analysis (GGUF)

A ComfyUI node in 🤖 GGUF-VLM/🖼️ Vision Models with 10 inputs and 1 output.

By walke2019·Created 9 months ago·Updated 8 months ago· 29
🖼️ Local Image Analysis (GGUF)
  • model
  • image
  • video
  • context
promptDescribe this image in detail.
max_tokens512
temperature0.7
top_p0.90
top_k40
seed0
system_promptYou are a helpful assistant that describes images and videos accurately and in detail.
Category🤖 GGUF-VLM/🖼️ Vision Models

Inputs (10)

NameTypeDefaultDescription
modelVISION_MODEL视觉语言模型配置
promptSTRINGDescribe this image in detail.用户提示词
max_tokensINT5121–4096最大生成 token 数
temperatureFLOAT0.70–2温度参数
top_pFLOAT0.900–1Top-p 采样
top_kINT400–100Top-k 采样
seedINT00–18446744073709550000随机种子
imageoptIMAGE输入图像(与视频二选一)
videooptIMAGE输入视频帧序列(与图像二选一)
system_promptoptSTRINGYou are a helpful assistant that describes images and videos accurately and in detail.系统提示词(可自定义模型行为)

Outputs (1)

NameTypeDescription
contextSTRING