ComfyUI Node: Emu 3.5 VQA
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.
Category
Emu3.5
Inputs
model EMU_MODEL
tokenizer EMU_TOKENIZER
vq_model EMU_VQ
image IMAGE
task_type
- caption
- describe
- analyze
- ocr
- question
- compare
- custom
max_tokens INT
question STRING
temperature FLOAT
image_resolution
- 256x256
- 384x384
- 512x512
image2 IMAGE
Outputs
STRING
Extension: Emu35-Comfyui-Nodes
ComfyUI integration for BAAI's Emu3.5 multimodal models for text-to-image generation and multimodal understanding. (Description by CC)
Authored by EricRollei
Looking for a different node?
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.