ComfyUI-ZoeyTool
ComfyUI多功能工具包 - 图像/视频处理、翻译、提示词生成
Zoey Tool
<a name="english"></a>
English
Multi-functional ComfyUI custom nodes plugin for image/video processing, translation, prompt generation, and more.
Installation
cd ComfyUI/custom_nodes
git clone https://github.com/liangzoey/comfyui-ZoeyTool.git
cd comfyui-ZoeyTool
pip install -r requirements.txt
Nodes
🖼️ Image Processing
| Node | Description | |------|-------------| | True Size Image Loader | Load images from folder at original resolution, sequential batch output, natural sorting, alpha channel as mask output | | Batch Image Saver | Save images + text files in batch with custom index, digit count, overwrite mode | | Mask Bounding Box Drawer | Draw bounding boxes from mask with HTML5 color picker, opacity/fill/width controls, preset colors, light type + prompt output, behind-subject compositing, optional rembg background removal | | Batch Image Cropper | Batch crop images (left/right/top/bottom), overwrite mode, filename preservation | | Light Handle Control | Interactive lighting direction with draggable handle, circular gradient mask, lighting prompt output (16 light types), behind-subject compositing, multiple handle shapes | | Text Overlay | Overlay text on images with custom or system fonts, adjustable size/color/position | | Multi-function Image Editor | Flip, rotate, split, edge detect, blur, sharpen, threshold, color invert, grayscale, perspective warp, blend, stylize | | Image Edit Prompt Generator | Auto-generate edit prompts for text/object/style/background editing, virtual try-on, object add/remove | | Frame Crop/Outpaint | Interactive frame-based cropping and outpainting with fill color, real-time canvas preview | | Multi-layer Canvas | 5-layer image compositing with drag reposition, scroll scale, rotation handle, and horizontal/vertical flip |
🎥 Video Processing
| Node | Description | |------|-------------| | Batch Video Loader | Load video file list from directory, multi-format (mp4/avi/mov), natural sort, limit control | | Video Batch Processor | Batch process videos (format, frame rate), auto-install dependencies, skip existing | | Smart Video Saver | Save processed videos with date folders and metadata |
📝 Text Tools
| Node | Description | |------|-------------| | Pure Translator | Multi-engine translation: Helsinki-NLP models / Baidu API / Google API. Supports ZH/EN/JA/KO | | Hunyuan Translator (HY-MT1.5) | Tencent Hunyuan MT local deployment, 20+ languages, auto-download model cache, terminology & context support | | Wan2.2 Prompt Generator | Cinematic video prompt generator with 16 control dimensions: subject, scene, action, lighting, composition, lens, style |
🔧 Other
| Node | Description | |------|-------------| | Multi File Batch Renamer | Batch rename files with regex find/replace, natural sort, custom start index | | ZOEYTextEncodeQwenImageEditPlus | Qwen image edit encoder with multi-reference images for Qwen2-VL | | VR 360° Preview | 360° equirectangular panorama VR viewer powered by Pannellum | | System Monitor | Live CPU/RAM/VRAM/GPU temperature & utilization dashboard via embedded HTTP server, with background VRAM auto-cleanup | | MiniMax H3 Reference-to-Video (@) | Wraps official MiniMaxH3ReferenceToVideo with Seedance-style @ reference tags (@P1/@V1/@A1/@C1/@L1) auto-expanded to native tags, plus a storyboard director panel and a permanent global asset library |
MiniMax H3 Usage
Wraps the official MiniMaxH3ReferenceToVideo (text + image/video/audio references → video). Type @ in the prompt to open a media picker with thumbnails; hover any @ thumbnail for a large preview.
@ Reference tags — auto-expand to native <Picture N>/<Video N>/<Audio N>:
| Tag | Meaning |
|-----|---------|
| @P1 | 1st connected reference image |
| @V1 | 1st connected reference video |
| @A1 | 1st connected reference audio |
| @C1 | storyboard character slot (director mode) |
| @L1 | global asset library entry |
🧰 Global Asset Library (permanent, cross-workflow): the 素材库 widget on the node lets you upload local images (character/prop/scene) and audio files to disk (input/zoey_library/). Characters support an appearance note and an optional voice audio. Reference any saved entry with @L1…; the node auto-loads the file and injects it into generation (no LoadImage node needed), and auto-writes purpose annotations such as <Picture K> 是{name}的人物参考(锁定脸和服装)。外貌:…,音色参考 <Audio N>.
🎬 Director panel (故事板/分镜): enable director_mode to build multi-shot storyboards instead of a single prompt. Per shot:
- Camera & shot-size preset buttons (push/pull/pan/truck/arc/POV/static/… + close-up/medium/wide/long/…)
- Dialogue rows (speaker S1–S5 + language + original text) compiled to
(S1) says: <d>[lang] text</d>; global speaker list for voice descriptions - Character slots
@C(assign a reference image → auto<Picture K> 是{name}的人物参考(锁定脸和服装)declaration) - Transition presets, duration with one-click auto-allocate by dialogue length, timeline start timestamps, shot reorder/duplicate, shot templates (product ad / character story / transition rhythm / music MV)
- Reference-purpose quick annotation, sound/music fields (
overall_soundscape/non_diegetic_music), cross-shot consistency toggle, and a live compiled-prompt preview
Highlights
- Mask Bounding Box Drawer: HTML5 color picker, opacity/fill/width controls, light type + prompt output, optional rembg background removal
- Translation: Pure Translator (Helsinki-NLP / Baidu / Google) + Hunyuan Translator (20+ languages, auto model cache)
- Video Pipeline: Batch load → process → save complete workflow
- Image Editor: 15+ operations including flip, rotate, blur, edge detect, sharpen, blend, stylize
- MiniMax H3: Seedance-style @ reference tags, permanent global asset library (local upload + auto-inject), and a full storyboard director panel
- System Monitor: live CPU/GPU/VRAM dashboard with background VRAM auto-cleanup
Dependencies
| Package | Required | Notes | |---------|----------|-------| | torch, Pillow, numpy | Yes | Core | | opencv-python-headless | Yes | Image/video ops | | transformers | For Hunyuan | Translator model | | rembg | Optional | Mask node BG removal | | av / PyAV | Optional | Video processor | | psutil | Optional | System monitor node |
License
MIT
<a name="chinese"></a>
中文
ComfyUI 多功能工具插件,提供图像/视频处理、翻译、提示词生成等实用节点。
安装
cd ComfyUI/custom_nodes
git clone https://github.com/liangzoey/comfyui-ZoeyTool.git
cd comfyui-ZoeyTool
pip install -r requirements.txt
节点列表
🖼️ 图像处理
| 节点 | 功能 | |------|------| | 真尺寸图像加载器 | 批量逐张加载文件夹中的图像,保持原始尺寸,自然排序,输出 Alpha 通道作为遮罩 | | 智能图像存储器 | 批量保存图像 + 文本文件,自定义序号位数、起始索引、覆盖模式 | | 遮罩边界框绘制 | 根据遮罩绘制矩形框,弹出调色板选色,透明度/填充/线宽调节,预设配色、灯光类型与提示词输出、主体后合成,可选 rembg 背景移除 | | 图像批量裁剪器 | 批量裁剪图片(左/右/上/下自由裁剪),覆盖模式,保留原文件名 | | 灯光手柄控制 | 交互式灯光方向控制,拖拽手柄生成圆形渐变遮罩与灯光提示词(16 种灯光类型),支持主体后合成、多种手柄形状 | | 文本叠加 | 在图像上叠加文字,支持自定义或系统字体,可调字号、颜色、位置 | | 多功能图像编辑器 | 翻转、旋转、分割、边缘检测、模糊、锐化、二值化、颜色反转、灰度化、透视变换、融合、风格化 | | 图像编辑提示词生成器 | 自动生成编辑提示词,支持文字/对象/风格/背景编辑、虚拟试穿、对象添加/移除 | | 框选裁剪/外扩 | 交互式框选裁剪与画布外扩,实时预览,自定义填充色 | | 多功能画布 | 5 层图像叠加合成,拖拽移动、滚轮缩放、旋转手柄、水平/垂直翻转 |
🎥 视频处理
| 节点 | 功能 | |------|------| | 批量视频加载器 | 从目录加载视频文件列表,支持多格式(mp4/avi/mov),自然排序 | | 视频批处理器 | 批量处理视频(格式转换、帧率调整),自动安装依赖,跳过已处理文件 | | 智能视频存储器 | 保存处理后视频,支持按日期分类、元数据写入 |
📝 文本工具
| 节点 | 功能 | |------|------| | 纯净翻译器 | 多引擎翻译:Helsinki-NLP 内置模型 / 百度API / 谷歌API,支持中/英/日/韩 | | 混元翻译器 (HY-MT1.5) | 腾讯混元翻译模型本地部署,20+ 语言,自动下载缓存,术语干预和上下文翻译 | | Wan2.2提示词生成器 | 影视级视频提示词生成,16 个控制维度:主体、场景、动作、光源、构图、镜头、风格 |
🔧 其他
| 节点 | 功能 | |------|------| | 多文件批量重命名 | 批量重命名文件,正则查找替换,自然排序,自定义起始序号 | | ZOEYTextEncodeQwenImageEditPlus | Qwen 图像编辑编码器,支持多张参考图 | | VR 360° 预览 | 360° 全景图 VR 预览,基于 Pannellum 全屏查看器 | | 系统监控 | 实时监控 CPU/内存/显存/GPU 温度与利用率,内置 HTTP 看板,后台自动清理显存 | | MiniMax H3 参考转视频 (@) | 包装官方 MiniMaxH3ReferenceToVideo,支持 Seedance 风格 @P1/@V1/@A1/@C1/@L1 引用语法自动展开为原生标签,附带导演台分镜面板与永久全局素材库 |
MiniMax H3 使用方法
包装官方 MiniMaxH3ReferenceToVideo(文字 + 图片/视频/音频参考 → 视频)。提示词里输入 @ 弹出带缩略图的素材选择器;鼠标悬停任意 @ 缩略图可看大图。
@ 引用标签 —— 自动展开为原生 <Picture N>/<Video N>/<Audio N>:
| 标签 | 含义 |
|------|------|
| @P1 | 第 1 张已连接的参考图 |
| @V1 | 第 1 段已连接的参考视频 |
| @A1 | 第 1 段已连接的参考音频 |
| @C1 | 导演台角色槽(导演模式) |
| @L1 | 全局素材库条目 |
🧰 永久全局素材库(跨工作流):节点上的「素材库」控件可本地上传图片(角色/道具/场景)与音频文件到磁盘(input/zoey_library/)。角色支持外貌备注与可选语音音频。提示词里用 @L1… 调用,节点自动加载文件并注入生成(无需再连 LoadImage 节点),并自动补用途标注,如 <Picture K> 是{name}的人物参考(锁定脸和服装)。外貌:…,音色参考 <Audio N>。
🎬 导演台(分镜/故事板):开启 director_mode 用多镜头分镜替代单条提示词。每镜支持:
- 景别/运镜快捷按钮(特写/近景/中景/全景/远景 + 推/拉/摇/横移/环绕/POV/静态…)
- 对白行(说话人 S1–S5 + 语言 + 台词原文)自动拼成
(S1) says: <d>[语言] 原文</d>;全局说话人列表写音色描述 - 角色槽
@C(分配参考图 → 自动生成<Picture K> 是{name}的人物参考(锁定脸和服装)声明) - 转场预设、时长 + 一键按台词分配时长、镜头起始时间轴、镜头复制/上下移、镜头模板(产品广告/角色剧情/转场节奏/音乐MV)
- 参考用途快捷标注、音效/配乐字段(
overall_soundscape/non_diegetic_music)、跨镜一致开关、实时编译预览
功能亮点
- 遮罩边界框绘制:弹出式调色板、透明度/填充/线宽调节、灯光类型与提示词输出、可选 rembg 背景移除
- 翻译:纯净翻译器(Helsinki-NLP/百度/谷歌) + 混元翻译器(20+语言、自动缓存)
- 视频流水线:批量加载→处理→保存完整工作流
- 图像编辑器:15+ 种操作(翻转、旋转、模糊、边缘检测、锐化、融合、风格化)
- MiniMax H3:Seedance 风格 @ 引用标签、永久全局素材库(本地上传 + 自动注入)、完整的导演台分镜面板
- 系统监控:实时 CPU/GPU/显存看板,后台自动清理显存
依赖
| 包 | 必需 | 说明 | |----|------|------| | torch, Pillow, numpy | 是 | 核心 | | opencv-python-headless | 是 | 图像/视频处理 | | transformers | 混元翻译器 | 翻译模型 | | rembg | 可选 | 遮罩节点背景移除 | | av / PyAV | 可选 | 视频处理器 | | psutil | 可选 | 系统监控节点 |
许可
MIT