Nodes/ComfyUI-llama-multimodal/[llama.cpp] Thinking / Reasoning Profile
ComfyUI Node

[llama.cpp] Thinking / Reasoning Profile

Controls thinking/reasoning mode, effort, and token budget for supported model templates. A zero token limit applies no separate reasoning budget.

By craftingmod·Created 2 months ago·Updated 5 days ago· 3
[llama.cpp] Thinking / Reasoning Profile
    • reasoning
    ◄reasoning_modeauto►
    ◄reasoning_effortauto►
    ◄max_reasoning_tokens0►
    ◄preserve_thinkingfalse►
    Categoryllama_cpp/profile

    Inputs (4)

    NameTypeDefaultDescription
    reasoning_modeCOMBOautoauto leaves template controls untouched; off and on explicitly disable or enable reasoning.
    reasoning_effortCOMBOautoUsed only when reasoning_mode is on.
    max_reasoning_tokensINT00–655360 applies no separate reasoning limit. Reasoning and final output still share Generate's max_tokens.
    preserve_thinkingBOOLEANfalsePreserve thinking/reasoning content in the chat history passed to the model.Currently, Qwen3.5+ only and may not work.

    Outputs (1)

    NameTypeDescription
    reasoningOLLAMA_IMAGE_LIST_LLAMA_CPP_REASONING_CONFIG—