ComfyUI Node: TurboQuant KV Patch
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.
Category
TurboQuant
Inputs
model MODEL
enabled BOOLEAN
Outputs
MODEL
Extension: ComfyUI-TurboQuant
TQ3 KV cache compression for ComfyUI reducing attention KV cache VRAM by ~4.5x using 3-bit Lloyd-Max quantization with Fast Walsh-Hadamard Transform decorrelation.
Authored by Scottcjn
Looking for a different node?
More nodes in ComfyUI-TurboQuant
Run ComfyUI workflows without the setup
No installs, no CUDA version roulette, no GPU sitting idle on your bill. Bring a workflow and run it in the browser.