Nodes/ComfyUI-FeiFei/Image Captioner (API: edit config.json)
ComfyUI Node

Image Captioner (API: edit config.json)

A ComfyUI node in FeiFei with 6 inputs and 3 outputs.

By aifeifei798·Created 13 days ago·Updated 2 days ago· 2
Image Captioner (API: edit config.json)
    • chinese
    • english
    • thinking
    ◄image▾►
    ◄instructionLook at this image carefully and output ONLY one JSON object, nothing else: {"chinese": "describe the image content, subject, action, scene, lighting and style in detail in Chinese", "english": "English image-generation prompt, comma-separated tags and quality words, directly usable in Stable Diffusion or Qwen-Image, e.g. '1girl, ... , masterpiece, best quality'"}►
    ◄temperature0.70►
    ◄max_tokens1024►
    ◄max_side1024►
    ◄thinking_modeOurs (8-step)►
    CategoryFeiFei

    Inputs (6)

    NameTypeDefaultDescription
    imageCOMBO1 options: example.png
    instructionSTRINGLook at this image carefully and output ONLY one JSON object, nothing else: {"chinese": "describe the image content, subject, action, scene, lighting and style in detail in Chinese", "english": "English image-generation prompt, comma-separated tags and quality words, directly usable in Stable Diffusion or Qwen-Image, e.g. '1girl, ... , masterpiece, best quality'"}—
    temperatureFLOAT0.700.1–1.5—
    max_tokensINT102464–8192—
    max_sideINT1024256–4096—
    thinking_modeCOMBOOurs (8-step)3 options: Ours (8-step), Model native, Both

    Outputs (3)

    NameTypeDescription
    chineseSTRING—
    englishSTRING—
    thinkingSTRING—