ComfyUI Node: Qwen-VL Model Loader
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.
Category
Qwen-VL
Inputs
model_name
- (none)
quantization
- 4-bit (VRAM-friendly)
- 8-bit (Balanced)
- None (FP16)
attention_mode
- auto
- sage
- flash_attention_2
- sdpa
device
- auto
- cuda
- cpu
- mps
use_compile BOOLEAN
Outputs
QWENVL_MODEL
Extension: ComfyUI Qwen-VL LoRA
Load Qwen-VL and Qwen3-VL models locally and apply PEFT LoRA adapters for enhanced image captioning. Supports 4-bit and 8-bit quantization (BitsAndBytes), FP8 pre-quantized models, Flash Attention 2, SageAttention, and torch.compile. Includes a LoRA strength slider and configurable captioning prompts.
Authored by Dangocan
Looking for a different node?
More nodes in ComfyUI Qwen-VL LoRA
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.