ComfyUI Node
Depth Anything 3
A ComfyUI node in š§ŖAILab/Geometry with 7 inputs and 1 output.
Depth Anything 3
- images
- IMAGE
āmodeldepth_anything_3_small.safetensorsāŗ
āresolution512āŗ
ānormalizationmin_maxāŗ
ācolormapgrayāŗ
āunload_modelfalseāŗ
āweight_dtypedefaultāŗ
Categoryš§ŖAILab/Geometry
Inputs (7)
| Name | Type | Default | Description |
|---|---|---|---|
| images | IMAGE | Input video frames or single image. Connect from Load Video (e.g. VHS Video Helper Suite) or Load Image. | |
| model | COMBO | depth_anything_3_small.safetensors | Depth Anything 3 model weight (auto-downloads if missing): - small: Fastest, lowest VRAM (~0.08B params, great for long videos & quick testing) - base: Great balance between speed and precision (~0.12B params) - mono_large: Highest visual depth detail (~0.35B params, best choice for AI video generation guiding) - metric_large: Real-world metric depth in meters (~0.35B params) |
| resolution | INT | 512128ā2560 | Model internal processing resolution (longest side, multiple of 14): - 504: Recommended default (fast, low VRAM) - 756 / 1008: Sharper edges & finer depth details (requires more VRAM) The output is always automatically upsampled back to the original input resolution. |
| normalization | COMBO | min_max | Depth range normalization mode: - min_max: Standard 0 to 1 range (near=white 1.0, far=black 0.0). Required for AI video generation (Wan/CogVideoX/Hunyuan/SVD) and ControlNet! - v2_style: Adaptive contrast balance with automatic sky clipping - raw: Unscaled numerical depth values (preserves metric units for 3D/VFX workflows) |
| colormap | COMBO | gray | Output color format: - gray: Standard grayscale (Required for AI Video Generation, ControlNet, and Depth-to-Video) - inferno: High-contrast orange/purple thermal heatmap (best for visual human inspection) - turbo: Smooth rainbow colormap (visual human inspection) |
| unload_model | BOOLEAN | false | Unload model from VRAM immediately after execution to free maximum GPU memory for downstream models (e.g. Wan 2.1, CogVideoX, HunyuanVideo). |
| weight_dtype | COMBO | default | Computation precision: - default: Uses model native weights - fp16: Recommended for Nvidia GPUs (saves ~50% VRAM, runs faster) - bf16: Recommended for RTX 30xx/40xx & newer GPUs - fp32: Full precision (uses highest VRAM) |
Outputs (1)
| Name | Type | Description |
|---|---|---|
| IMAGE | IMAGE | ā |