Nodes/ComfyUI-ArchAi3d-Qwen/ArchAi3D_Qwen_Encoder_V3
ComfyUI Node

ArchAi3D_Qwen_Encoder_V3

A ComfyUI node in ArchAi3d/Qwen/Encoders with 21 inputs and 4 outputs.

By amir84ferdos·Created 10 months ago·Updated 4 months ago· 65
ArchAi3D_Qwen_Encoder_V3
  • clip
  • vae
  • image1_vl
  • image2_vl
  • image3_vl
  • image1_latent
  • image2_latent
  • image3_latent
  • conditioning
  • latent
  • formatted_prompt
  • recommended_cfg
prompt
system_prompt
conditioning_balance
conditioning_balance_override
manual_context_strength1.00
manual_user_strength1.00
image1_labelImage 1
image2_labelImage 2
image3_labelImage 3
image1_latent_strength1.00
image2_latent_strength1.00
image3_latent_strength1.00
debug_modefalse
CategoryArchAi3d/Qwen/Encoders

Inputs (21)

NameTypeDefaultDescription
clipCLIPQwen-VL CLIP model for encoding text and vision tokens
promptSTRINGText prompt (vision tokens inserted automatically in ChatML format)
system_promptSTRINGOptional system prompt (wrapped in ChatML <|im_start|>system block)
conditioning_balanceCOMBOV3 PRESET: Choose conditioning balance (Image-Dominant → Text-Dominant). Works great with ConditioningAverage!
conditioning_balance_overrideSTRINGOPTIONAL: Connect ⚖️ Conditioning Balance node here to override the preset above. Leave empty to use the dropdown.
manual_context_strengthFLOAT1.000–3CUSTOM ONLY: Manual context strength (only used when preset = Custom). Extended range: 0.0-3.0 for extreme conditioning control
manual_user_strengthFLOAT1.000–3CUSTOM ONLY: Manual user strength (only used when preset = Custom). Extended range: 0.0-3.0 for extreme conditioning control
image1_labelSTRINGImage 1Custom label for Image 1
image2_labelSTRINGImage 2Custom label for Image 2
image3_labelSTRINGImage 3Custom label for Image 3
image1_latent_strengthFLOAT1.000–2Image1 latent strength (1.0=normal, <1.0=weaker, >1.0=stronger)
image2_latent_strengthFLOAT1.000–2Image2 latent strength (1.0=normal, <1.0=weaker, >1.0=stronger)
image3_latent_strengthFLOAT1.000–2Image3 latent strength (1.0=normal, <1.0=weaker, >1.0=stronger)
debug_modeBOOLEANfalseEnable console logging (shows preset values, strengths, shapes)
vaeoptVAEVAE for encoding reference latents (required if using latent images)
image1_vloptIMAGEImage 1 for vision encoder (RGB only, expects correct size)
image2_vloptIMAGEImage 2 for vision encoder (RGB only, expects correct size)
image3_vloptIMAGEImage 3 for vision encoder (RGB only, expects correct size)
image1_latentoptIMAGEImage 1 for reference latent (RGB only, expects correct size)
image2_latentoptIMAGEImage 2 for reference latent (RGB only, expects correct size)
image3_latentoptIMAGEImage 3 for reference latent (RGB only, expects correct size)

Outputs (4)

NameTypeDescription
conditioningCONDITIONINGText+vision embeddings with reference latents metadata attached
latentLATENTImage1 latent in standard format (for VAEDecode or other latent nodes)
formatted_promptSTRINGFinal ChatML-formatted prompt with vision tokens (for debugging)
recommended_cfgFLOAT⭐ NEW: Recommended CFG scale based on preset (2.5-5.5). Connect to KSampler's cfg parameter for optimal image/text balance!