Nodes/ComfyUI-QING/Qwen视觉丨API
ComfyUI Node

Qwen视觉丨API

A ComfyUI node in QING/API with 10 inputs and 3 outputs.

By sheengoa·Created 12 months ago·Updated 3 months ago· 16
Qwen视觉丨API
  • image
  • analysis_result
  • conversation_info
  • total_tokens
text_input请描述这张图片的内容。
platform阿里云百炼
modelqwen2.5-vl-72b-instruct
max_tokens2048
history5
temperature0.3
top_p0.80
image_qualityauto
clear_historyfalse
CategoryQING/API

Inputs (10)

NameTypeDefaultDescription
imageIMAGE输入要分析的图像
text_inputSTRING请描述这张图片的内容。输入要发送给Qwen视觉模型的文本问题,Qwen-VL擅长图像理解、文档分析、OCR识别和视觉推理
platformCOMBO阿里云百炼选择API服务提供商
modelCOMBOqwen2.5-vl-72b-instruct选择要使用的Qwen视觉模型 📋 阿里云百炼模型特点: 🔸 qwen3-vl-plus:新一代视觉模型,图像理解能力强,推荐首选 🔸 qwen3-vl-235b-a22b-instruct:大参数视觉模型,精度更高 🔸 qwen-vl-max-latest:最新旗舰视觉模型,功能最全面 🔸 qwen2.5-vl-72b-instruct:经典大模型版本,稳定可靠 📋 硅基流动模型特点: 🔸 qwen2.5-vl-72b-instruct:开源版本,性价比高 💡 Qwen-VL系列在图像理解、文档OCR、图表分析、视觉推理方面表现优异
max_tokensINT20481–32768模型生成文本时最多能使用的token数量
historyINT51–25保持的历史对话轮数
temperatureoptFLOAT0.30–2控制生成文本的随机性
top_poptFLOAT0.800–1控制生成文本的多样性
image_qualityoptCOMBOauto图像处理质量:auto(自动选择), low(低质量,速度快), high(高质量,精度高)
clear_historyoptBOOLEANfalse是否清除历史对话记录

Outputs (3)

NameTypeDescription
analysis_resultSTRING
conversation_infoSTRING
total_tokensINT