ComfyUI Node
Youtu-VL (GGUF)
A ComfyUI node in 🧪AILab/YoutuVL with 7 inputs and 1 output.
Youtu-VL (GGUF)
- image
- text
◄modelYoutu-VL-4B-Instruct-GGUF-Q8►
◄preset_prompt🖼️ Describe Image►
◄custom_prompt►
◄max_tokens512►
◄keep_model_loadedtrue►
◄seed1►
Category🧪AILab/YoutuVL
Inputs (7)
| Name | Type | Default | Description |
|---|---|---|---|
| model | COMBO | Youtu-VL-4B-Instruct-GGUF-Q8 | Select the GGUF quantized model. |
| preset_prompt | COMBO | 🖼️ Describe Image | 6 options: 🖼️ Describe Image, 📝 Detailed Description, 🔍 Analyze Elements, 🏷️ Generate Tags, 📄 OCR Text, 🎨 Art Style Analysis |
| custom_prompt | STRING | — | |
| max_tokens | INT | 51264–4096 | Maximum number of new tokens to generate. |
| keep_model_loaded | BOOLEAN | true | Keep model loaded for faster subsequent inference. |
| seed | INT | 11–4294967295 | — |
| imageopt | IMAGE | — |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| text | STRING | — |