ComfyUI Node
OneVision Caption Folder
A ComfyUI node in LLaVA-OneVision with 10 inputs and 1 output.
OneVision Caption Folder
- llava_model
- STRING
◄folder_path—►
◄promptYou are AI captioning tool, you caption images in very elaborate detail without referring to the image as 'the image', the results should be useful for image model training purposes. You focus on the composition, style and action any possible subject is performing. You don't make assumptions or try to tell a story. You also describe the background of the image separately. Caption this image:►
◄max_tokens512►
◄keep_model_loadedtrue►
◄temperature0.20►
◄seed1►
◄max_image_size1024►
◄prefix►
◄suffix►
CategoryLLaVA-OneVision
Inputs (10)
| Name | Type | Default | Description |
|---|---|---|---|
| llava_model | LLAVAMODEL | — | |
| folder_path | STRING | — | |
| prompt | STRING | You are AI captioning tool, you caption images in very elaborate detail without referring to the image as 'the image', the results should be useful for image model training purposes. You focus on the composition, style and action any possible subject is performing. You don't make assumptions or try to tell a story. You also describe the background of the image separately. Caption this image: | — |
| max_tokens | INT | 5121–8192 | — |
| keep_model_loaded | BOOLEAN | true | — |
| temperature | FLOAT | 0.200–1 | — |
| seed | INT | 11–18446744073709550000 | — |
| max_image_size | INT | 1024256–8192 | — |
| prefix | STRING | — | |
| suffix | STRING | — |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| STRING | STRING | — |