Extensions/RunningHub MiniMax H3
ComfyUI Extension

RunningHub MiniMax H3

RunningHub MiniMax-H3 audio-video custom nodes for ComfyUI (direct, in-process)

By RH-RunningHub·Created 28 days ago·Updated 17 days ago· 1
RH-RunningHub/ComfyUI-RH-MiniMax-H3
Nodes33
On cloudLocal install
CategoryRunningHub/MiniMax H3/decode, RunningHub/MiniMax H3/loaders
Stars1
Updated17 days ago

Nodes (33)

RunningHub MiniMax H3 Decode Video + Audio (Legacy)

Where MiniMax H3 Clips Actually Become Watchable

RunningHub/MiniMax H3/decode
RunningHub MiniMax H3 Model Loader (Direct) (Legacy)

The Legacy MiniMax H3 DiT Loader That Still Works Fine

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Qwen3-VL Loader (Direct) (Legacy)

The Qwen3-VL Loader Behind Every MiniMax H3 Prompt

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Dual VAE Loader (Direct) (Legacy)

Where MiniMax H3 Keeps Video and Audio Separate

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Dual Sigma Sampler (Legacy)

Where MiniMax H3 Video and Audio Finally Get Denoised

RunningHub/MiniMax H3/sampling
RunningHub MiniMax H3 Empty AV Latent (Legacy)

MiniMax H3's Blank Canvas

RunningHub/MiniMax H3/latent
RunningHub MiniMax H3 Encode Video → AV Latent (Legacy)

The Door to MiniMax H3 Video-to-Audio

RunningHub/MiniMax H3/latent
RunningHub MiniMax H3 FL2VA Encode (Legacy)

Turning Keyframes and a Prompt Into H3 Conditioning

RunningHub/MiniMax H3/fl2va
RunningHub MiniMax H3 FL2VA First / First+Last (Legacy)

The Simplest Way to Steer a MiniMax H3 Clip

RunningHub/MiniMax H3/fl2va
RunningHub MiniMax H3 FL2VA Last Only (Legacy)

The Weird but Useful FL2VA Sibling

RunningHub/MiniMax H3/fl2va
RunningHub MiniMax H3 FL2VA Model Loader (Direct) (Legacy)

Pick This When You Know You're Doing Keyframes

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 FL2VA Target (Legacy)

Choosing the Canvas Size and Length of Your H3 Clip

RunningHub/MiniMax H3/fl2va
RunningHub MiniMax H3 FL2VA Qwen3-VL Loader (Direct) (Legacy)

The FL2VA-Fixed Qwen3-VL Encoder Loader — a Legacy Node That Just Works

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 FL2VA Dual VAE Loader (Direct) (Legacy)

The FL2VA Dual VAE Loader — Video VAE Plus Audio VAE, Pinned to One Partition

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Frame Rate (Experimental) (Legacy)

Trying to Make MiniMax H3 Do 30 or 60 FPS

RunningHub/MiniMax H3/sampling
RunningHub MiniMax H3 Model Loader

The One Loader You Actually Need for MiniMax H3 in ComfyUI

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Ref2VA Audio Reference (Legacy)

Matching a Voice or a Sound With MiniMax H3

RunningHub/MiniMax H3/ref2va
RunningHub MiniMax H3 Ref2VA Encode (Legacy)

Fusing References and Prompt Into H3 Conditioning

RunningHub/MiniMax H3/ref2va
RunningHub MiniMax H3 Ref2VA Image Reference (Legacy)

Handing MiniMax H3 a Face (or a Product) to Keep

RunningHub/MiniMax H3/ref2va
RunningHub MiniMax H3 Ref2VA Model Loader (Direct) (Legacy)

The Ref2VA DiT Loader for MiniMax H3's Multimodal Reference Model

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Ref2VA Target (Legacy)

The legacy Ref2VA target node — where H3 decides its size from your references

RunningHub/MiniMax H3/ref2va
RunningHub MiniMax H3 Ref2VA Qwen3-VL Loader (Direct) (Legacy)

The legacy 'Direct' Qwen3-VL loader, locked to the Ref2VA partition

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Ref2VA Dual VAE Loader (Direct) (Legacy)

The legacy 'Direct' dual VAE loader, Ref2VA edition

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Ref2VA Video Reference (Legacy)

Feed MiniMax H3 a video reference, audio track included if you want it

RunningHub/MiniMax H3/ref2va
RunningHub MiniMax H3 Video Gen (References)

Images, clips, and audio in one autogrowing node

RunningHub/MiniMax H3
RunningHub MiniMax H3 Sampler Config (Experimental) (Legacy)

Sparse attention and step aborts

RunningHub/MiniMax H3/sampling
RunningHub MiniMax H3 Separate AV Latent (Legacy)

Split MiniMax H3's audio-video latent into its two halves

RunningHub/MiniMax H3/latent
RunningHub MiniMax H3 T2VA Target (Legacy)

The legacy node that decided your MiniMax H3 canvas — and why you can ignore it now

RunningHub/MiniMax H3/conditioning
RunningHub MiniMax H3 T2VA Text Encode (Legacy)

The legacy node that turns your prompt into MiniMax H3 conditioning

RunningHub/MiniMax H3/conditioning
RunningHub MiniMax H3 Qwen3-VL Loader

The Qwen3-VL loader that powers every MiniMax H3 prompt

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Legacy Unsupported Conditioning (Migration Error) (Legacy)

This node is a dead end on purpose — it's a migration error

RunningHub/MiniMax H3/conditioning
RunningHub MiniMax H3 Dual VAE Loader

MiniMax H3 needs two VAEs, and this node loads both at once

RunningHub/MiniMax H3/loaders
RunningHub MiniMax H3 Video Gen (Text / Keyframes)

Text, keyframes, or dubbing in one box

RunningHub/MiniMax H3
Readme

ComfyUI-RH-MiniMax-H3

RunningHub China RunningHub International English 简体中文

License

Native MiniMax-H3 audio-video generation nodes for ComfyUI. Model components run inside the ComfyUI process without an SGLang server or Diffusers pipeline. Based on the official MiniMax-H3 project.

✨ Features

  • T2VA, FL2VA, Ref2VA (including audio-only references), and video-to-audio generation
  • A focused UI with two generation nodes and three model loaders
  • INT8 ConvRot weights, optional turbo LoRA, offload, and attention backends
  • Attention backends: auto, sdpa, sage, and ck (Comfy Kitchen INT8, same kernel as --use-ck-attention)
  • Example workflows for text, keyframe, multimodal-reference, and V2A tasks

🛠️ Installation

cd ComfyUI/custom_nodes
git clone https://github.com/RH-RunningHub/ComfyUI-RH-MiniMax-H3.git
pip install -r ComfyUI-RH-MiniMax-H3/requirements.txt

Restart ComfyUI after installation.

📦 Model Download & Installation

Use the complete converted INT8 ConvRot model package.

| Priority | Source | Purpose | |---|---|---| | 1 | Hugging Face INT8 ConvRot | Preferred converted weights | | 2 | ModelScope INT8 ConvRot | Preferred China mirror |

The converted repositories contain the same complete bundle. Run one of the following commands from the ComfyUI root directory. The download is about 95 GiB, so keep at least 110 GiB of free disk space.

Model Directory Structure

All models should be placed in ComfyUI/models/MiniMax-H3-INT8-CONVROT/ with the following structure:

ComfyUI/
└── models/
    └── MiniMax-H3-INT8-CONVROT/
        ├── MiniMax-H3-FL2VA-int8_convrot.safetensors
        ├── MiniMax-H3-Ref2VA-int8_convrot.safetensors
        ├── qwen3-vl-32b-int8_convrot.safetensors
        ├── MiniMax-H3-video_vae.safetensors
        ├── MiniMax-H3-audio_vae.safetensors
        ├── minimax_h3_fl2v_turbo_4step_v0.1.safetensors
        ├── minimax_h3_fl2v_turbo_4step_v1.0_768p_bf16.safetensors
        ├── minimax_h3_fl2v_turbo_8step_v1.0_bf16.safetensors
        ├── minimax_h3_ref2v_turbo_4step_v0.1_bf16.safetensors
        ├── FL2VA/
        │   ├── model_index.json
        │   ├── transformer/config.json
        │   ├── text_encoder/config.json
        │   ├── tokenizer/
        │   ├── processor/
        │   ├── video_vae/config.json
        │   ├── video_vae/source/config.json
        │   └── audio_vae/config.json
        └── Ref2VA/
            └── ... same configuration layout

The plugin automatically detects the complete converted bundle. The legacy ComfyUI/models/MiniMax-H3 directory remains supported.

Download Methods

Method 1: Download from Hugging Face (Recommended)

cd /path/to/ComfyUI
python3 -m pip install -U huggingface_hub
hf download Gluttony10/MiniMax-H3-INT8-CONVROT \
  --local-dir ./models/MiniMax-H3-INT8-CONVROT

Re-run the same command to resume or update an interrupted download. On a high-bandwidth machine with at least 64 GiB RAM, prefix the hf download command with HF_XET_HIGH_PERFORMANCE=1 for maximum throughput.

Method 2: Download from ModelScope (For China users)

cd /path/to/ComfyUI
python3 -m pip install -U modelscope
modelscope download --model Gluttony10/MiniMax-H3-INT8-CONVROT \
  --local_dir ./models/MiniMax-H3-INT8-CONVROT

Method 3: Manual Download

| Model | Link | Description | |---|---|---| | Hugging Face bundle | Gluttony10/MiniMax-H3-INT8-CONVROT | Complete INT8 ConvRot package, VAEs, and turbo LoRAs | | ModelScope bundle | Gluttony10/MiniMax-H3-INT8-CONVROT | Same package for China users |

Optional Turbo LoRA

Select one turbo LoRA in the model loader when you want fewer sampling steps. Leave the LoRA empty to run the base converted weights.

| File | Typical use | |---|---| | minimax_h3_fl2v_turbo_4step_v0.1.safetensors | FL2VA / T2VA 4-step turbo | | minimax_h3_fl2v_turbo_4step_v1.0_768p_bf16.safetensors | FL2VA / T2VA 4-step turbo for 768p | | minimax_h3_fl2v_turbo_8step_v1.0_bf16.safetensors | FL2VA / T2VA 8-step turbo | | minimax_h3_ref2v_turbo_4step_v0.1_bf16.safetensors | Ref2VA 4-step turbo |

Model weights are not covered by this repository's Apache-2.0 license. Review the upstream model terms before use.

🚀 Usage

The main nodes are under RunningHub/MiniMax H3. Both generation nodes output frames, audio, and av_latent. Start with the bundled workflows:

Older workflows can be migrated with:

python3 tools/migrate_workflow.py old_workflow.json --in-place

Legacy granular nodes remain registered so existing workflows still load and run, but they are deprecated and hidden from search, the node tree, and slot-drag suggestions.

📝 Node Reference

| Node | Purpose | |---|---| | RHMiniMaxH3ModelLoader | Load the FL2VA / Ref2VA DiT and optional LoRA | | RHMiniMaxH3TextEncoderLoader | Load the Qwen3-VL text encoder | | RHMiniMaxH3VAELoader | Load the video and audio VAEs | | RHMiniMaxH3VideoGen | T2VA / FL2VA / V2A from text, keyframes, or a source video | | RHMiniMaxH3RefGen | Ref2VA with ordered image, video, and audio references; audio-only is allowed |

attention_backend on the loaders and generation nodes:

| Value | Behavior | |---|---| | auto | Follow ComfyUI's current optimized attention | | sdpa | PyTorch SDPA | | sage | SageAttention when available | | ck | Comfy Kitchen INT8 attention (same kernel as --use-ck-attention; errors if unavailable) |

Requirements

  • ComfyUI 0.27 or newer (0.28+ recommended)
  • A ComfyUI-compatible CUDA build of PyTorch, Triton, and comfy-kitchen
  • ffmpeg and ffprobe for Ref2VA video/audio references
  • Python packages listed in requirements.txt
  • Sufficient RAM, VRAM, and fast model storage for MiniMax-H3

📄 License

Plugin code is licensed under Apache-2.0. Model weights use their upstream licenses. See NOTICE.md for attribution and third-party notices.

🔗 Links

RunningHub China RunningHub International

🙏 Acknowledgements

This project is based on MiniMax-H3, developed by MiniMax-AI.