返回项目目录
wildminder

wildminder

awesome-minimax-H3

Awesome MiniMax-H3

视觉 / 图像
Stars
243
Forks
9
Watchers
243
Issues
0

README

项目介绍

82841 bytes

Awesome MiniMax-H3

A curated list of models, text encoders, quants, and tools for the MiniMax-H3 omni-modal video generation model.

[![Telegram][telegram-shield]][telegram-url] [![X][x-shield]][x-url]
Table of Contents * [Models](#models) * [Checkpoints](#checkpoints) * [Quantized Models](#quants) * [GGUF](#gguf) * [Fine-tuned Checkpoints](#finetunes) * [Text Encoders](#text-encoder) * [Separated Components](#components) * [VAE (Video & Audio)](#components-vae) * [Tiny Autoencoder (TAE)](#tae) * [Image VAE (Mamad8)](#cliproj) * [Clip Projection (ClipProj)](#cliproj) * [LoRA](#lora) * [Styles](#lora) * [Turbo (Acceleration LoRA)](#lora) * [Experimental / Other](#lora) * [ComfyUI Nodes](#nodes) * [Custom Node Collections](#nodes) * [Special Stuff](#nodes) * [Guides & Tutorials](#guides) * [Workflow & Technical Notes](#wf) * [ComfyUI](#wf-comfyui)

Intro

▓ Models

MiniMax-H3 is a general-purpose, omni-modal generative system by MiniMaxAI. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. The model has two variants: FL2VA (first-and-last-frame mode) and Ref2VA (omni-reference mode).

▣ Checkpoints

Official and ComfyUI-repackaged model files.

Variant Name Precision Size Download
FL2VA minimax_h3_fl2va bf16 61.73 GB
FL2VA minimax_h3_fl2va int8 31.70 GB
FL2VA minimax_h3_fl2va_pruned bf16 37.46 GB
FL2VA minimax_h3_fl2va_pruned fp8 19.52 GB
FL2VA minimax_h3_fl2va_pruned int8 19.53 GB
Ref2VA minimax_h3_ref2va bf16 61.73 GB
Ref2VA minimax_h3_ref2va int8 31.70 GB
Ref2VA minimax_h3_ref2va_pruned bf16 37.46 GB
Ref2VA minimax_h3_ref2va_pruned fp8 19.52 GB
Ref2VA minimax_h3_ref2va_pruned int8 19.53 GB

Model Variants: - H3-Base-FL2VA (First-and-last-frame mode): Supports zero, one, or two input images. No image input = T2V; one image = first/last-frame-to-video; two images = first-and-last-frame-to-video. - H3-Base-Ref2VA (Omni-reference mode): Supports multi-modal reference inputs — up to 9 images, 3 video clips (2–15s each), 3 audio clips, max 12 files total.

▣ Turbo (Acceleration LoRA)

4-step audio-video generation LoRAs — render joint video + synchronized stereo audio in 4 sampling steps instead of ~20 (~5× speedup). Early prototype; comfort zone for sharpness is 6–8 steps. The lightx2v distil (top row) is the shared base for most ComfyUI conversions; for pruned checkpoints use the ComfyUI-converted variants below. The original larryvrh LoRA targets the full (non-pruned) FL2VA checkpoint and needs the ComfyUI-MiniMax-H3-Turbo sampler node.

Variant Steps Pruned / Full Precision Size Download
fl2v v0.1 4 Full bf16 1.29 GB
fl2v v1.0 768p 4 Full bf16 1.29 GB
fl2v v1.0 768p · comfyui 4 Full bf16 1.82 GB
fl2v v1.0 8 Full bf16 1.29 GB
fl2v v1.0 · comfyui 8 Full bf16 1.82 GB
lightx2v v0.1 4 Full bf16 1.82 GB
lightx2v v0.1 · resized 4 Full bf16 300 MB
fl2v 4 Full bf16 744 MB
fl2v ema 4 Full bf16 744 MB
fl2v ckpt500 4 Full bf16 744 MB
fl2v ema ckpt500 4 Full bf16 744 MB
fl2v ckpt850 4 Full bf16 744 MB
fl2v ema ckpt850 4 Full bf16 744 MB
fl2v v4 step600 4 Full bf16 744 MB
fl2v v4 step600 ema 4 Full bf16 744 MB
fl2v pruned 4 Pruned bf16 592 MB
fl2v pruned ema 4 Pruned bf16 592 MB
fl2v pruned ckpt500 4 Pruned bf16 592 MB
fl2v pruned ema ckpt500 4 Pruned bf16 592 MB
fl2v pruned ckpt850 4 Pruned bf16 592 MB
fl2v pruned ema ckpt850 4 Pruned bf16 592 MB
fl2v pruned v4 step600 4 Pruned bf16 592 MB
fl2v pruned v4 step600 ema 4 Pruned bf16 592 MB
fl2v v1.0 768p · resized 4 Pruned bf16 298 MB
fl2v v1.0 · resized 8 Pruned bf16 327 MB
fl2v pruned ckpt500 V1 4 Pruned bf16 592 MB
fl2v pruned ckpt600 V4 4 Pruned bf16 592 MB
fl2v pruned ckpt600 ema V4 4 Pruned bf16 592 MB
fl2v pruned ckpt850 V1 4 Pruned bf16 592 MB
fl2v 4 Full bf16 717 MB
fl2v step 100 8 NFE Full bf16 738 MB
fl2v step 200 8 NFE Full bf16 738 MB
fl2v step 300 8 NFE Full bf16 738 MB
fl2v 4-step acceleration · ConvRot · ⚠️ needs dual-clock sampler or 8–10 steps 4 Full int8 779.9 MB
fl2v 4-step acceleration ema · ConvRot 4 Full int8 779.9 MB
fl2v v4 step600 (T8-convert) · ConvRot 4 Full int8 779.9 MB
lightx2v v0.1 · alpha8 T8-convert · ConvRot · ⚠️ needs dual-clock sampler or 8–10 steps 4 Full int8 1.96 GB
fl2v 10ErosMax test4 · 4-step curveproj1025 (T8) · ConvRot · ⚠️ needs dual-clock sampler or 8–10 steps 4 Pruned int8 794.9 MB
fl2v 10ErosMax test4 · 4-step curveproj1025 4 Pruned int8 794.9 MB
fl2v 10ErosMax test4 · 8-step v1.0 · ConvRot 8 Pruned int8 1.96 GB
fl2v CMF · full 4 Full Q4TP (CMF) 25.20 GB
fl2v CMF · FL2VA 4 Full Q4TP (CMF) 25.70 GB
fl2v CMF · FL2VA (smaller) 4 Full Q2TP (CMF) 20.12 GB
fl2v v1.0 768p · ConvRot · needs ComfyUI-LoraInt8Loader 4 Full int8 991 MB
fl2v v1.0 · ConvRot · needs ComfyUI-LoraInt8Loader 8 Full int8 991 MB
lightx2v v0.1 · int8 · ConvRot · needs ComfyUI-LoraInt8Loader 4 Full int8 991 MB

larryvrh also publishes experimental training checkpoints (11 .bin files: step 149/490/729/850/922, v2 step 298, v3 step 300, v4 step 150/600, v5 step 600; 7.26–10.17 GB) — see the repo.

══════════════════════════════════

▣ Quantized Models

Unified quantization tables for FL2VA and Ref2VA. The Pruned column marks whether the checkpoint is AdaLN-pruned (smaller, ComfyUI-only). The Method column identifies the quantization scheme. Multiple sources for the same quant are separated by .

Key: ConvRot = ConvRotation INT8/INT4 quantization · Lean = selective BF16 island retention · DT-sQKV = Dynamic-Time separate-QKV (patch required) · W4A8 = 4-bit weight / 8-bit activation · GGUF = llama.cpp GGUF format · NF4 = bitsandbytes 4-bit · OrbitQuant = native W4A4 packed path · Hybrid = partial NVFP4 layers on Blackwell.

Items marked ⚠️ require a ComfyUI core patch — they do not load in unmodified ComfyUI.

FL2VA — Unified Quantization Table | Pruned | Precision | Method | Size | Download | | :---: | :---: | :--- | :---: | :--- | | | ![bf16][badge-bf16] | BF16 | 61.73 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_fl2va_bf16.safetensors) | | | ![bf16][badge-bf16] | Hybrid (fl2va base + ref2va adaln b15-49) | 20.97 GB | [![][gh-smhfacct]](https://huggingface.co/smhfacct/Minimax-H3-fl2va-ref2va-hybrid-models/resolve/main/minimax_h3_hybrid_fl2va_ref2va_b15-49.safetensors) | | | ![bf16][badge-bf16] | Hybrid (fl2va base + ref2va adaln b20-49) | 20.97 GB | [![][gh-smhfacct]](https://huggingface.co/smhfacct/Minimax-H3-fl2va-ref2va-hybrid-models/resolve/main/minimax_h3_hybrid_fl2va_ref2va_b20-49.safetensors) | | | ![bf16][badge-bf16] | Hybrid (fl2va base + ref2va adaln b25-49) | 20.97 GB | [![][gh-smhfacct]](https://huggingface.co/smhfacct/Minimax-H3-fl2va-ref2va-hybrid-models/resolve/main/minimax_h3_hybrid_fl2va_ref2va_b25-49.safetensors) | | | ![bf16][badge-bf16] | Hybrid (fl2va base + ref2va adaln b30-49) | 20.97 GB | [![][gh-smhfacct]](https://huggingface.co/smhfacct/Minimax-H3-fl2va-ref2va-hybrid-models/resolve/main/minimax_h3_hybrid_fl2va_ref2va_b30-49.safetensors) | | | ![int8][badge-int8] | ConvRot | 31.70 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_fl2va_int8_convrot.safetensors) | | | ![fp8][badge-fp8] | FP8 E4M3FN | 43.78 GB | [![][gh-rzgar]](https://huggingface.co/rzgar/minimax_h3_fl2va_fp8_e4m3fn/resolve/main/minimax_h3_fl2va_fp8_e4m3fn.safetensors) | | | ![mxfp8][badge-mxfp8] | MXFP8 | 44.34 GB | [![][gh-rzgar]](https://huggingface.co/rzgar/minimax_h3_fl2va_fp8_e4m3fn/resolve/main/minimax_h3_fl2va_mxfp8.safetensors) | | | ![fp8][badge-fp8] | FP8 + FP16 attn | 26.70 GB | [![][gh-rzgar]](https://huggingface.co/rzgar/minimax_h3_fl2va_fp8_e4m3fn/resolve/main/minimax_h3_fl2va_fp16attn_fp8.safetensors) | | | ![int8][badge-int8] | ConvRot Lean | 21.91 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/FL2VA/MiniMax-H3_FL2VA-INT8-ConvRot-HQ.safetensors) | | | ![int8][badge-int8] | ConvRot | 20.94 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/FL2VA/MiniMax-H3_FL2VA-INT8-ConvRot.safetensors) | | | ![int8][badge-int8] | ConvRot Lite | 20.33 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/FL2VA/MiniMax-H3_FL2VA-INT8-ConvRot-Lite.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 13.60 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/FL2VA/MiniMax-H3_FL2VA-NVFP4-HQ.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 10.86 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/FL2VA/MiniMax-H3_FL2VA-NVFP4.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 32.05 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_fl2va_nvfp4.safetensors) | | | ![int4][badge-int4] | NF4 | 15.98 GB | [![][gh-DiffSynth-Studio]](https://huggingface.co/DiffSynth-Studio/MiniMax-H3-NF4/resolve/main/minimax-h3-fl2va-nf4.safetensors) | | | | OrbitQuant W4A4 | 17.03 GB | [![][gh-WaveCut]](https://huggingface.co/WaveCut/MiniMax-H3-OrbitQuant-W4A4/resolve/main/transformer/diffusion_pytorch_model-00001-of-00005.safetensors) | | | ![int8][badge-int8] | ⚠️ DT-sQKV ConvRot | 21.00 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-DynTime-sQKV/resolve/main/FL2VA/MiniMax-H3_FL2VA-DT-sQKV-INT8-ConvRot.safetensors) | | | ![int8][badge-int8] | ⚠️ DT-sQKV ConvRot Lean | 27.99 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-DynTime-sQKV/resolve/main/FL2VA/MiniMax-H3_FL2VA-DT-sQKV-INT8-ConvRot-HQ.safetensors) | | ✓ | ![bf16][badge-bf16] | BF16 | 37.46 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_fl2va_pruned_bf16.safetensors) | | ✓ | ![fp8][badge-fp8] | FP8 scaled | 19.52 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_fl2va_pruned_fp8_scaled.safetensors) | | ✓ | ![int8][badge-int8] | ConvRot | 19.53 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_fl2va_pruned_int8_convrot.safetensors) ┊ [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_FL2VA_pruned_int8_convrot.safetensors) | | ✓ | ![nvfp4][badge-nvfp4] | NVFP4 | 18.69 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_fl2va_pruned_nvfp4.safetensors) | | ✓ | ![nvfp4][badge-nvfp4] | NVFP4 + ConvRot INT8 | 18.69 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_fl2va_pruned_nvfp4_convrot_int8.safetensors) | | ✓ | ![nvfp4][badge-nvfp4] | NVFP4 | 11.67 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_FL2VA_pruned_nvfp4.safetensors) | | ✓ | ![int4][badge-int4] | Mixed INT4/INT8 ConvRot | 14.81 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_FL2VA_pruned_mixed_int4_int8_convrot.safetensors) ┊ [![][gh-tsolful]](https://huggingface.co/tsolful/Minimax_H3_INT4MixedConvRot/resolve/main/minimax_h3_fl2va_pruned_INT4BQ.safetensors) | | ✓ | ![int4][badge-int4] | Mixed INT4/INT8 ConvRot Lean | 17.27 GB | [![][gh-tsolful]](https://huggingface.co/tsolful/Minimax_H3_INT4MixedConvRot/resolve/main/minimax_h3_fl2va_pruned_INT4Q.safetensors) | | ✓ | ![int4][badge-int4] | INT4 ConvRot | 15.67 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_fl2va_pruned_int4_convrot_simple.safetensors) | | ✓ | ![int4][badge-int4] | Mixed INT4/INT8 ConvRot | 18.92 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_fl2va_pruned_mixed_int4_int8_convrot_simple.safetensors) | | ✓ | ![int4][badge-int4] | W4A8 ConvRot | 11.68 GB | [![][gh-AX1Y2JP]](https://huggingface.co/AX1Y2JP/MiniMax-H3-W4A8-ConvRot/resolve/main/minimax_h3_fl2va_pruned_symw4a8convrot.safetensors) ┊ [![][gh-Kijai]](https://huggingface.co/Kijai/MiniMax-H3-experimental/resolve/main/minimax_h3_fl2va_pruned_w4a8_mixed.safetensors) ┊ [![][gh-Winnougan]](https://huggingface.co/Winnougan/MiniMax-H3-INT4_Convrot_ComfyUI/resolve/main/minimax_h3_fl2va_pruned-w4a8_convrot_pruned.safetensors) | *GGUF quants — see [GGUF section](#gguf) below.* Ref2VA — Unified Quantization Table | Pruned | Precision | Method | Size | Download | | :---: | :---: | :--- | :---: | :--- | | | ![bf16][badge-bf16] | BF16 | 61.73 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_ref2va_bf16.safetensors) | | | ![int8][badge-int8] | ConvRot | 31.70 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_ref2va_int8_convrot.safetensors) ┊ [![][gh-t8star]](https://huggingface.co/t8star/minimax_h3_ref2va_patchin_hf102/resolve/main/minimax_h3_ref2va_patchin_hf102_T8.safetensors) | | | ![int8][badge-int8] | ConvRot Lean | 21.91 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-INT8-ConvRot-HQ.safetensors) | | | ![int8][badge-int8] | ConvRot | 20.94 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-INT8-ConvRot.safetensors) | | | ![int8][badge-int8] | ConvRot Lite | 20.33 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-INT8-ConvRot-Lite.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 13.60 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-NVFP4-HQ.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 10.86 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-ComfyUI-Quants/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-NVFP4.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 32.05 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_ref2va_nvfp4.safetensors) | | | ![nvfp4][badge-nvfp4] | NVFP4 | 22.76 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_Ref2VA_nvfp4_mixed.safetensors) | | | ![int4][badge-int4] | NF4 | 15.98 GB | [![][gh-DiffSynth-Studio]](https://huggingface.co/DiffSynth-Studio/MiniMax-H3-NF4/resolve/main/minimax-h3-ref2va-nf4.safetensors) | | | | OrbitQuant W4A4 | 17.03 GB | [![][gh-WaveCut]](https://huggingface.co/WaveCut/MiniMax-H3-OrbitQuant-W4A4/resolve/main/transformer_ref/diffusion_pytorch_model-00001-of-00005.safetensors) | | | ![nvfp4][badge-nvfp4] | Hybrid NVFP4 (FFN-only) | 16.38 GB | [![][gh-abakanai]](https://huggingface.co/abakanai/Minimax_h3_hybrid/resolve/main/minimax_h3_ref2va_pruned_hybrid_ffn_nvfp4_blackwell.safetensors) | | | ![nvfp4][badge-nvfp4] | Hybrid NVFP4 (QKV+FFN) | 14.03 GB | [![][gh-abakanai]](https://huggingface.co/abakanai/Minimax_h3_hybrid/resolve/main/minimax_h3_ref2va_pruned_hybrid_nvfp4_blackwell.safetensors) | | | ![int8][badge-int8] | ⚠️ DT-sQKV ConvRot | 21.00 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-DynTime-sQKV/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-DT-sQKV-INT8-ConvRot.safetensors) | | | ![int8][badge-int8] | ⚠️ DT-sQKV ConvRot Lean | 27.99 GB | [![][gh-DmitryDB]](https://huggingface.co/DmitryDB/MiniMax-H3-DynTime-sQKV/resolve/main/Ref2VA/MiniMax-H3_Ref2VA-DT-sQKV-INT8-ConvRot-HQ.safetensors) | | ✓ | ![bf16][badge-bf16] | BF16 | 37.46 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_ref2va_pruned_bf16.safetensors) | | ✓ | ![fp8][badge-fp8] | FP8 scaled | 19.52 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_ref2va_pruned_fp8_scaled.safetensors) | | ✓ | ![int8][badge-int8] | ConvRot | 19.53 GB | [![][gh-Comfy--Org]](https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/diffusion_models/minimax_h3_ref2va_pruned_int8_convrot.safetensors) ┊ [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_Ref2VA_pruned_int8_convrot.safetensors) | | ✓ | ![nvfp4][badge-nvfp4] | NVFP4 | 18.69 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_ref2va_pruned_nvfp4.safetensors) | | ✓ | ![nvfp4][badge-nvfp4] | NVFP4 + ConvRot INT8 | 18.69 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_ref2va_pruned_nvfp4_convrot_int8.safetensors) | | ✓ | ![nvfp4][badge-nvfp4] | NVFP4 | 11.67 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_Ref2VA_pruned_nvfp4.safetensors) | | ✓ | ![int4][badge-int4] | Mixed INT4/INT8 ConvRot | 14.06 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/Minimax-H3-nvfp4-INT4-INT8-Convrot/resolve/main/MiniMax_H3_Ref2VA_pruned_mixed_int4_int8_convrot.safetensors) ┊ [![][gh-tsolful]](https://huggingface.co/tsolful/Minimax_H3_INT4MixedConvRot/resolve/main/minimax_h3_ref2va_pruned_INT4BQ.safetensors) | | ✓ | ![int4][badge-int4] | Mixed INT4/INT8 ConvRot Lean | 17.18 GB | [![][gh-tsolful]](https://huggingface.co/tsolful/Minimax_H3_INT4MixedConvRot/resolve/main/minimax_h3_ref2va_pruned_INT4Q.safetensors) | | ✓ | ![int4][badge-int4] | INT4 ConvRot | 15.67 GB | [![][gh-rockerBOO]](https://huggingface.co/rockerBOO/minimax-h3-nvfp4/resolve/main/minimax_h3_ref2va_pruned_int4_convrot_simple.safetensors) | | ✓ | ![int4][badge-int4] | W4A8 ConvRot | 11.68 GB | [![][gh-AX1Y2JP]](https://huggingface.co/AX1Y2JP/MiniMax-H3-W4A8-ConvRot/resolve/main/minimax_h3_ref2va_pruned_symw4a8convrot.safetensors) ┊ [![][gh-Kijai]](https://huggingface.co/Kijai/MiniMax-H3-experimental/resolve/main/minimax_h3_ref2va_pruned_w4a8_mixed.safetensors) ┊ [![][gh-Winnougan]](https://huggingface.co/Winnougan/MiniMax-H3-INT4_Convrot_ComfyUI/resolve/main/minimax_h3_ref2va_pruned-w4a8_convrot_pruned.safetensors) | *GGUF quants — see [GGUF section](#gguf) below.*

· · · · · · · · · · · · · ·

GGUF Quantized Models

GGUF quants for use with stable-diffusion.cpp, ComfyUI, and Unsloth. Non-pruned sources: Abiray/MiniMax-H3-GGUF, vantagewithai/MiniMax-H3-comfyUI-GGUF, realrebelai/MiniMax-H3_GGUFs. Pruned sources: unsloth/MiniMax-H3-GGUF, MarxistLeninist/MiniMax-H3-FL2VA-Pruned-IQ1-GGUF.

FL2VA GGUF | Pruned | Quant | Size | Download | | :---: | :---: | :---: | :--- | | | ![Q2_K][badge-Q2_K] | 17.42 GB | [![][gh-realrebelai]](https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/resolve/main/MiniMax-H3-FL2VA-Q2_K-(Mixed_Precision).gguf) | | | ![Q3_K_M][badge-Q3_K_M] | 14.50 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q3_K_M.gguf) ┊ [![][gh-realrebelai]](https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/resolve/main/MiniMax-H3-FL2VA-Q3_K_M.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q3_K_M.gguf) | | |![Q3_K_S][badge-Q3_K_S] | 14.50 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q3_K_S.gguf) | | | ![Q4_0][badge-Q4_0] | 17.36 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q4_0.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q4_0.gguf) | | | ![Q4_1][badge-Q4_1] | 20.41 GB | [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q4_1.gguf) | | | ![Q4_K_M][badge-Q4_K_M] | 18.50 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q4_K_M.gguf) ┊ [![][gh-realrebelai]](https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/resolve/main/MiniMax-H3-FL2VA-Q4_K_M.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q4_K_M.gguf) | | | ![Q4_K_S][badge-Q4_K_S] | 18.49 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q4_K_S.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q4_K_S.gguf) | | | ![Q5_0][badge-Q5_0] | 21.21 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q5_0.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q5_0.gguf) | | | ![Q5_1][badge-Q5_1] | 24.17 GB | [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q5_1.gguf) | | | ![Q5_K_M][badge-Q5_K_M] | 22.25 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q5_K_M.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q5_K_M.gguf) | | | ![Q5_K_S][badge-Q5_K_S] | 22.25 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q5_K_S.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q5_K_S.gguf) | | | ![Q6_K][badge-Q6_K] | 26.28 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q6_K.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q6_K.gguf) | | | ![Q8_0][badge-Q8_0] | 33.56 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-FL2VA-Q8_0.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/fl2va/minimax_h3_fl2va-Q8_0.gguf) | | ✓ | ![Q2_K][badge-Q2_K] | 6.26 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-Q2_K.gguf) | | ✓ | ![Q3_K_M][badge-Q3_K_M] | 8.16 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-Q3_K.gguf) | | ✓ | ![Q4_K_M][badge-Q4_K_M] | 10.64 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-Q4_K.gguf) | | ✓ | ![Q5_0][badge-Q5_0] | 12.97 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-Q5_0.gguf) | | ✓ | ![Q6_K][badge-Q6_K] | 15.45 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-Q6_K.gguf) | | ✓ | ![Q8_0][badge-Q8_0] | 19.97 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-Q8_0.gguf) | | ✓ | ![UD-Q2_K_XL][badge-UD-Q2_K_XL] | 7.51 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-UD-Q2_K_XL.gguf) | | ✓ | ![UD-Q3_K_XL][badge-UD-Q3_K_XL] | 8.90 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_fl2va_pruned-UD-Q3_K_XL.gguf) | | ✓ | ![IQ1_S][badge-IQ1_S] | 3.78 GB | [![][gh-MarxistLeninist]](https://huggingface.co/MarxistLeninist/MiniMax-H3-FL2VA-Pruned-IQ1-GGUF/resolve/main/minimax_h3_fl2va_pruned-IQ1_S.gguf) | | ✓ | ![IQ1_M][badge-IQ1_M] | 4.22 GB | [![][gh-MarxistLeninist]](https://huggingface.co/MarxistLeninist/MiniMax-H3-FL2VA-Pruned-IQ1-GGUF/resolve/main/minimax_h3_fl2va_pruned-IQ1_M.gguf) | Ref2VA GGUF | Pruned | Quant | Size | Download | | :---: | :---: | :---: | :--- | | | ![Q3_K_M][badge-Q3_K_M] | 14.50 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q3_K_M.gguf) ┊ [![][gh-realrebelai]](https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/resolve/main/MiniMax-H3-REF2VA-Q3_K_M.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q3_K_M.gguf) | | | ![Q3_K_S][badge-Q3_K_S] | 14.50 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q3_K_S.gguf) | | | ![Q4_0][badge-Q4_0] | 17.36 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q4_0.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q4_0.gguf) | | | ![Q4_1][badge-Q4_1] | 20.41 GB | [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q4_1.gguf) | | | ![Q4_K_M][badge-Q4_K_M] | 18.49 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q4_K_M.gguf) ┊ [![][gh-realrebelai]](https://huggingface.co/realrebelai/MiniMax-H3_GGUFs/resolve/main/MiniMax-H3-REF2VA-Q4_K_M.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q4_K_M.gguf) | | | ![Q4_K_S][badge-Q4_K_S] | 18.49 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q4_K_S.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q4_K_S.gguf) | | | ![Q5_0][badge-Q5_0] | 21.21 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q5_0.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q5_0.gguf) | | | ![Q5_1][badge-Q5_1] | 24.17 GB | [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q5_1.gguf) | | | ![Q5_K_M][badge-Q5_K_M] | 22.25 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q5_K_M.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q5_K_M.gguf) | | | ![Q5_K_S][badge-Q5_K_S] | 22.25 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q5_K_S.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q5_K_S.gguf) | | | ![Q6_K][badge-Q6_K] | 26.28 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q6_K.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q6_K.gguf) | | | ![Q8_0][badge-Q8_0] | 33.56 GB | [![][gh-Abiray]](https://huggingface.co/Abiray/MiniMax-H3-GGUF/resolve/main/unet/MiniMax-H3-Ref2VA-Q8_0.gguf) ┊ [![][gh-vantagewithai]](https://huggingface.co/vantagewithai/MiniMax-H3-comfyUI-GGUF/resolve/main/ref2va/minimax_h3_ref2va-Q8_0.gguf) | | ✓ | ![Q2_K][badge-Q2_K] | 6.22 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_ref2va_pruned-Q2_K.gguf) | | ✓ | ![Q3_K_M][badge-Q3_K_M] | 8.12 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_ref2va_pruned-Q3_K.gguf) | | ✓ | ![Q4_K_M][badge-Q4_K_M] | 10.60 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_ref2va_pruned-Q4_K.gguf) | | ✓ | ![Q5_0][badge-Q5_0] | 12.94 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_ref2va_pruned-Q5_0.gguf) | | ✓ | ![Q6_K][badge-Q6_K] | 15.42 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_ref2va_pruned-Q6_K.gguf) | | ✓ | ![Q8_0][badge-Q8_0] | 19.94 GB | [![][gh-unsloth]](https://huggingface.co/unsloth/MiniMax-H3-GGUF/resolve/main/minimax_h3_ref2va_pruned-Q8_0.gguf) |

· · · · · · · · · · · · · ·

Fine-tuned Checkpoints

Stock-compatible quants for the 10Eros_Max fine-tune of MiniMax-H3. Fine-tuned QKV weights in blocks 0–31 preserved alongside tested quantization layouts. No custom node or ComfyUI core patch required. (DmitryDB/MiniMax-H3-10Eros-Max-Quants)

Variant Precision Method Size Download
FL2VA 10Eros int8 ConvRot Lean 21.91 GB
FL2VA 10Eros int8 ConvRot 20.94 GB
FL2VA 10Eros nvfp4 NVFP4 13.60 GB
FL2VA 10Eros nvfp4 NVFP4 10.86 GB

Patch-required FL2VA for the 10Eros_Max fine-tune. DT-sQKV edition (DmitryDB/MiniMax-H3-10Eros-Max-DT-sQKV):

Variant Precision Method Size Download
FL2VA 10Eros int8 ⚠️ DT-sQKV ConvRot 21.00 GB

Notes

  • t8star Ref2VA patchin HF 1.02 — experimental weight modification (not a quant): +2% on 2×2 spatial HF patch in the video-input projection. Tests showed weak HF agent gain; "oily/waxy" look not confirmed removed. Repo. (31.70 GB, INT8 ConvRot, listed in the Ref2VA table above with *(patchin)* label.)
  • DmitryDB/MiniMax-H3-INT8-Lean-ConvRot is the same repo as DmitryDB/MiniMax-H3-ComfyUI-Quants (merged/rebranded by the author). Both names resolve to the same files.
  • DmitryDB/MiniMax-H3-INT8-Lean-ConvRot-Dynamic-Time-Separate-QKV is the same repo as DmitryDB/MiniMax-H3-DynTime-sQKV. Both names resolve to the same files.
  • Winnougan/MiniMax-H3-INT4_Convrot_ComfyUI also includes a matching quantized text encoder: qwen3vl_32b_minimax_h3-w4a8_convrot.safetensors.
  • Kijai/MiniMax-H3-experimental also includes an INT8 ConvRot video VAE: minimax_h3_video_vae_int8_convrot.safetensors (2.95 GB). See Components.
  • unsloth/MiniMax-H3-GGUF also includes Qwen3-VL text encoder GGUFs: Q2_K_M (12.2 GB) and Q4_K_M (17.0 GB).
  • DmitryDB/MiniMax-H3-ComfyUI-Quants also includes VAE files: Video VAE FP16 (4.85 GB) and Audio VAE FP32 (577 MB). See Components.
  • DiffSynth-Studio/MiniMax-H3-NF4 also includes TE, Video VAE, and Audio VAE NF4 quants. Requires DiffSynth-Studio; minimum 8 GB VRAM.
  • WaveCut/MiniMax-H3-OrbitQuant-W4A4 also includes quantized text encoder and FP32 VAE copies. Requires ComfyUI-OrbitQuant custom node. Workflow JSON.
  • DeepBeepMeep/MiniMax-H3 is a community repack bundling both FL2VA and Ref2VA in every precision/pruning combination: full bf16 (66.3 GB) and int8_convrot (34 GB); pruned bf16 (41.4 GB) and int8_convrot (22.1 GB); and pruned_rank8 bf16 (40.3 GB) and int8_convrot (21.1 GB). Also ships VAEs (video fp16 5.21 GB, video fp8mix 2.79 GB, audio fp32 605 MB), a Qwen3-VL-32B text encoder (nvfp4_awq + Q4_K_M GGUF), and SeedVR2 upscaler checkpoints. No license is stated — clarify usage rights before redistributing. Repo

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Text Encoders

MiniMax-H3 uses the Qwen3-VL-32B model as its text/vision conditioning encoder.

▣ Comfy-Org Optimized Encoders

Official and optimized versions for ComfyUI, repackaged by Comfy-Org.

Model Name Precision Size Download
qwen3vl_32b_minimax_h3 bf16 47.97 GB
qwen3vl_32b_minimax_h3 int8 25.28 GB
qwen3vl_32b_minimax_h3 nvfp4 14.61 GB

▣ Abiray GGUF Text Encoder

GGUF quantized text encoder, bundled with the Abiray/MiniMax-H3-GGUF repository.

Model Name Precision Size Download
qwen3vl_32b_minimax_h3 Q4_K_M 13.58 GB
qwen3vl_32b_minimax_h3 int4 13.93 GB
qwen3vl_32b_minimax_h3 nvfp4 25.28 GB

· · · · · · · · · · · · · ·

▣ Qwen3-VL-32B Ultra-Heretic (Uncensored)

Built from llmfan46/Qwen3-VL-32B-Instruct-ultra-uncensored-heretic by ethanfel. Includes a MiniMax-H3 conditioning encoder (language layers 0–49 + vision tower) and an optional prompt-enhancement tail (layers 50–63 + LM head). The "Heretic" lineage bypasses alignment/restriction layers in the text encoder so MiniMax-H3 receives the most faithful prompt embeddings.

Model Name Precision Size Download
qwen3vl_32b_heretic (conditioning encoder) int8 24.55 GB
qwen3vl_32b_heretic (generation tail 50–63) int8 7.09 GB

The generation tail is loaded temporarily by the ComfyUI-MiniMax-H3-Guide node for prompt enhancement, then unloaded. Requires the connected standard MiniMax-H3 CLIP (layers 0–49).

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Separated Components

Separated VAE files for MiniMax-H3. The video VAE and audio VAE are required for all generation workflows.

▣ VAE (Video & Audio)

Component Source Precision Size Download
Video VAE Comfy-Org fp16 4.85 GB
Audio VAE Comfy-Org fp32 577 MB
Video VAE dummy9996 fp8 2.60 GB
Audio VAE dummy9996 bf16 289 MB

FP8-mixed quantized VAE set by dummy9996 — smaller video VAE (2.60 GB, fp8) and audio VAE (289 MB, bf16) for low-VRAM workflows.

· · · · · · · · · · · · · ·

▣ Tiny Autoencoder (TAE)

Quickly trained 2D tiny VAE for MiniMax-H3 by Kijai. Not the greatest outcome, still beats latent2rgb for preview purposes. Currently only works with the ModelPreviewOverride node in ComfyUI-KJNodes.

Component Size Download
TAE (preview VAE) 9 MB

▣ Image VAE (Mamad8)

Experimental image-specialized MiniMax H3 VAE that decodes a single temporal latent (T=1) into one image. Merged H3 VAE checkpoint — no custom node required. For image workflows only; the image-tuned decoder materially regresses multi-frame video reconstruction, so keep the original H3 VAE for video.

Component Size Download
Single-image VAE (step 1597) 4.85 GB

· · · · · · · · · · · · · ·

▣ Clip Projection (ClipProj + Conditioning)

Learned linear projections to condition H3 from a smaller text encoder. Two families: (1) ClipProj — swap the large Qwen3-VL-32B for a 4B/8B one (text-encoder VRAM ~15.7 GB → 4.5 GB, no change to the diffusion model, VAE, or sampler), and (2) H3 Control — identity/zero matrices for a no-control baseline. Projection files are fp16, MIT-licensed. Requires the ComfyUI-ClipProj node; place files in ComfyUI/models/clip_projections/. Full variant matrix (4B/8B × base/MLP/celeb/celeb-MLP): repo.

Variant Encoder Size Download
ClipProj (base) Qwen3-VL 4B 52.5 MB
ClipProj (MLP) Qwen3-VL 4B 304 MB
ClipProj (celeb) Qwen3-VL 4B 52.5 MB
ClipProj (celeb-MLP) Qwen3-VL 4B 304 MB
ClipProj (base) Qwen3-VL 8B 84 MB
ClipProj (MLP) Qwen3-VL 8B 386 MB
ClipProj (celeb) Qwen3-VL 8B 84 MB
ClipProj (celeb-MLP) Qwen3-VL 8B 386 MB
H3 Control Identity 52.5 MB
H3 Control Zero 52.5 MB

Older h3_* filenames (with tap24 / CONDPROJ / int8convrot suffixes) have moved to obsolete/ — canonical names are now mmh3-*-ClipProj*.safetensors.

· · · · · · · · · · · · · ·

▣ Ref Patch (lihaoyun6)

fl2varef2va behavior patch by lihaoyun6. Extracts 112 specific keys shared between the ref2va and fl2va weights and stores their differences as a single patch, letting the lighter FL2VA checkpoint partially mimic Ref2VA output quality. Requires the ComfyUI-MiniMaxH3_Ref-Patch node to load. Apache-2.0.

Component Size Download
Ref Patch 148 MB

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ LoRA

▣ Styles

  • SexGod1979
  • PinkFluffyBunny - Pink fluffy bunny style LoRA in pruned + unpruned variants (rank 128/256/512). Maximum pink achieved at 0.5 strength on pruned int8 model. Alpha quality. (2.31 GB · pruned-v1 rank128)
  • PinkCherry - High-quality furry rabbits, rainbows, and cherry blossoms. No guardrails altered. Alpha v0.3 (pruned int8, 14 GB checkpoint). Iterated alpha 0.1→0.5.
  • NaughtyTimes - NSFW style LoRA for MiniMax-H3.

  • ssjenforcer191

  • Homelander - Character LoRA for The Boys' Homelander. Triggerword HeroHomelander (optionally append wearing red leather gloves). Experimental. (296 MB)

  • matlod/minimax-h3-turnaround - Contact-Sheet diffusion — one reference image + one instruction → five coherent, progressively rotated views of the same subject in a single pass. A character turnaround from one photo (~10 s at 512², ~57 s at 1024²). Uses H3's timeline as a slot axis. (60 MB each: 1024-cont/s600, 512/s1500, 512-instruct/s400)

  • EllaPriest45/MinimaxH3_Actions - Collection of NSFW action LoRAs for MiniMax-H3 (T2V/I2V/R2V). Includes motion-specific LoRAs with trigger words and strength recommendations. See the repo for the full list. (reference only)

  • fal/research-mini-max-h3-realism-people-lora - Realism LoRA for natural-looking people in everyday scenarios. Trained by fal on diverse photo data. (125 MB)

  • Inner-Reflections/MiniMax-H3-Looping-Sketch-Anime - Looping anime-style sketch LoRA. Hand-drawn 2D outlines, flat colors, white outline. Strength 0.75–1.25; pair with a Turbo LoRA for higher strength. (569 MB)

  • nikdevs/minimax-h3-loras - ⚠️ Contains explicit / NSFW content. Curated MiniMax-H3 LoRA collection (styles + characters). Browse at your own discretion; not enumerated with per-file downloads here.

  • DiffSynth-Studio/MiniMax-H3-LoRA-LineartAnime - Anime video line-art colorization — feeds a line-art video as a reference and generates fully colored anime output from it (Ref2VA video-reference workflow). Apache-2.0. (1.26 GB)

▣ Experimental / Other

  • bghira/minimax-h3-anyflow-wip - SimpleTuner WIP LoRA checkpoints (steps 200/300/400/500 + EMA). WIP research builds; not production-tuned.

  • ethanfel/MiniMax-H3-Pruned-Ref2VA-Delta-LoRAs-Experimental - Highly experimental, mechanically extracted adapters — randomized-SVD approximations of the weight difference between pruned FL2VA and Ref2VA checkpoints. Not trained as LoRAs, not generation-tested. Explore behavior transfer in either direction. (ranks 256/512/1024, BF16)

  • Kijai/MiniMax-H3-experimental loras - Experimental rank-256 BF16 LoRA capturing the FL2VA↔Ref2VA difference (same class as ethanfel's). No confirmed use case yet. (2.40 GB)

  • DIE2025/MiniMaxH3Loras - no description Three unnamed style LoRAs (B, Spicy, V) of equal size. No README; use at own discretion. (310 MB each)

Variant Size Download
MiniMaxB.safetensors 310 MB
MiniMaxSpicy.safetensors 310 MB
MiniMaxV.safetensors 310 MB
  • MATLOWAI/MiniMax-H3-Motion-Adapter - Motion adapter (pilot, r16) — a small rank-16 BF16 LoRA that improves the de-rope pass in ComfyUI-MAINodes on fast motion: reduces frame-by-frame advance/snap alternation and over-production, and transfers to both FL2VA and Ref2VA graphs (one file). Trained bf16 (rank 16, alpha 16). MIT for the adapter weights; base model use is under the MiniMax H3 Community License. Load with a stock LoraLoaderModelOnly at strength 1.0 on the de-rope pass only. (63 MB)

    ◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ ComfyUI Nodes

Node Author Category Description
MiniMax H3 Hybrid Cond kitsune123150 Conditioning Hybrid R2V + I2V conditioning in one payload. Outputs positive conditioning and AV latent with native audio.
ComfyUI-H3-Multishot jlucasmcrell Conditioning Multishot video+audio generation — N chained shots from one script, seam-clean master. Keyframes at any position, dual-format loader (safetensors + GGUF).
ComfyUI MiniMax H3 Director seesee75-commits Conditioning Timeline editor with storyboard — drag media onto tracks, trim on a ruler, write a prompt per shot. Live sampling preview, retakes, shot chaining.
ComfyUI MiniMax H3 Image Studio astropuzzo Conditioning Image-first nodes for T2I, I2I, and reference editing. Arbitrary frame counts, resolution up to 64 MP, automatic still-frame scoring.
ComfyUI-MiniMaxH3-Easy nkxx188 Conditioning One compact workflow for T2V, I2V, first/last-frame, and reference video. Unified multi-media input with @ references and inline dialogue blocks.
H3 Motion Context NikoDemon80 Conditioning Chain H3 clips so motion and sound keep going across the cut. Feed clip A's last frames + audio in; clip B picks up where A left off — same motion, same audio.
ComfyUI MiniMax H3 Motion Director j955229 Conditioning Multi-segment motion director combining AIMixer Director's timeline + Motion Context chaining. Reference control across N segments.
H3 Conditioning Cache HEEEeeeeN Conditioning Conditioning cache + batch generation suite for H3 drama/short-drama production. Caches conditioning across shots, batch-generates episodes unattended.
MAINodes matlowai Conditioning Contact-Sheet diffusion (five views from one reference) + Motion Lab (test-time de-roping of fast-motion smearing: backflips, sword arcs, reversals).
Fantastic MiniMax H3 Prompt Builder Adudeguyman Prompt Fillable prompt templates for every H3 mode with live guide-rule checking and a media loader that manages reference tags.
MiniMax-H3 Prompt Enhancer T8 T8mars Prompt Multimodal prompt enhancer calling doubao-seed-evolving. Analyzes text, images, and video together. Supports all H3 modes, strict/balanced/creative, CN/EN output.
MiniMaxH3 LatentUpscaler Tr1dae Upscaling Latent spatial upscaler for H3's NestedTensor AV latents. Re-noises video/audio for two-pass sampling, scales minimax_refs/minimax_keyframes conditioning.
ComfyUI Video Tiler maDcaDDie2000 Upscaling Memory-conscious video/image tiling with overlap tiles, gaps, and feather blending. Built for LTX 2.3 and MiniMax H3 tiled upscale workflows. Disk-backed mode for low-VRAM.
H3 Latent Upscaler (Mamad8) mamad8c Upscaling Moves a clean H3 video latent to a 2× larger spatial latent grid very quickly. Not a conventional upscaler — output looks softer than input; the point is to get a 2× grid ready for a second pass.
MiniMaxH3 Frame Infill red-polo Conditioning Experimental node to regenerate any frame interval of an existing H3 video. Patches ComfyUI's H3 internal implementation; pin your ComfyUI version.
ComfyUI-SolAttn_triton kijai Acceleration SolAttention Triton kernel for ComfyUI. Optimized attention computation for H3 and other Sol-Attn models.
ComfyUI-sol-attn Saganaki22 Acceleration Zero-copy Sol-Attn for SM89–SM120 with scheduled tau, graph preview, and feed-forward chunking. 1.14–1.44× vs SageAttention, −37% MLP peak VRAM on H3.
ComfyUI Spectrum MiniMax H3 xmarre Acceleration Spectral feature forecasting — skips selected transformer evaluations via Chebyshev ridge regression. Adaptive scheduling with native fallbacks.
Herrgotts-H3-Infinite-Continuation-Suite HerrgottMargott Conditioning Freeze-aware, keyframe-anchored MiniMax H3 video continuation for ComfyUI — injects the previous clip's video+audio latent context into the next FL2VA segment, auto-detects H3's frozen tail for a safe handover, and stitches with a 4-frame video crossfade + 15 ms audio de-click. Experimental community project (GPL-3.0).
ComfyUI-MiniMaxH3-Cache lihaoyun6 Acceleration EasyCache-style cache node for H3. Patches ComfyUI core to cache and reuse transformer block computations across timesteps.
MiniMax H3 Block Cache T8 T8mars Acceleration F1B0 block cache — computes Block 0 and reuses residual for Blocks 1–49 when audio/video are stable. Skips up to 49 of 50 blocks per step.
TE-Speed-MiniMaxH3-OSS HELPMEEADICE Acceleration Block-cache accelerator patching H3's 50-layer DiT loop. Reuses cached tail-block residuals when sigma delta is small. ~45% speedup at default settings.
MiniMaxH3 Dual-Clock Euler Sampler shuaixn Acceleration Dual-clock Euler sampler for the Turbo LoRA — fixes audio crackling/noise at 4-step generation by running video and audio on separate schedules.
minimax-h3-mlx mrbizarro Port Apple Silicon MLX port of the full H3 pipeline. AdaLN precompute drops 13B params at inference. Validated against the diffusers reference.
ComfyUI-ClipProj nicolab28 Port Swap a large text encoder for a small one via a learned linear projection. MiniMax H3 conditioning from 15.7 GB down to 5.2 GB. Proof of concept.
ComfyUI MiniMax H3 Contex Loop ethanfel Conditioning Turn one sampling body into a scene-by-scene production loop — each accepted scene carries motion + audio forward, saves a checkpoint, joins into final video without huge cumulative tensors.
ComfyUI MiniMax H3 LongMedia vizart-vj Acceleration Long single-pass video/audio generation with streamed Sol attention, compressed KV, adaptive VRAM guards, chunked MLP/final output. SAFE long-sequence optimizations for limited VRAM.
ComfyUI MiniMaxH3 Hybrid Loader scottmudge Port Load a checkpoint by merging selected tensor groups (e.g. adaln_proj only) from a ref2va overlay onto a fl2va base. Default preset preserves ref-conditioning pathway while keeping fl2va quality.
ComfyUI MiniMax H3 Legacy Audio Sampling starsFriday Acceleration Restores the v0.30.0 audio sampling behavior after upgrading to ComfyUI v0.31.0. One model-patch node — no source modification. Fixes regressed audio (background noise, stereo stability, HF artifacts).
ComfyUI-H3-FaceRefine Carasibana Face Refine Face-refinement node for MiniMax H3 outputs — repair/enhance faces in generated video frames.
ComfyUI-MiniMaxH3Mod Luisacaotica Conditioning No-training "RefMod" reference adapters for MiniMax H3 — compress reference images/videos into tiny .safetensors latent files reused like LoRAs without loading heavy references or training. Extract/Load/Apply nodes, folder and A/B-axis loaders, a standalone CLI, and strength/retention controls injected via the model's native conditioning path.
ComfyUI MiniMax H3 Extender tritant Conditioning Chains multiple H3 clips into one long continuous sequence, preserving motion, visual, and audio continuity. Combines Ref2VA conditioning, motion context, disk latent caching, dynamic image references (up to 9), audio reference support, per-clip prompt/seed/duration, clip validation, and seamless video/audio decoding with seam correction (H.264 / H.265 / FFV1 export).
ComfyUI ALLinONE MiniMaxH3 LeonQ8 Conditioning All-in-one MiniMax H3 node — T2V, I2V, R2V, audio drive (lip sync), keyframes, extend, chain (multi-clip continuation via H3 Motion Context), and an RTX/Seed2VR upscale hook in a single node. Ships searchable history, a library, and settings UI. Beta, GPL-3.0.
ComfyUI Qwen H3 Prompt chflame163 Prompt Generates H3 prompts inside ComfyUI with a local Qwen3.8-27B GGUF model (bundled llama-server, fully offline) plus the official MiniMax-H3 Skills. Routes modes (T2VA/I2VA/L2VA/FL2VA/Ref2VA) from image/video references, writes sound design, and supports think mode with per-reference image/video inputs.
OpenH3-IR ruashots Prompt Open-source, local implementation of MiniMax H3's Context-IR stage — compiles a plain-language sentence (with optional referenced media) into a structured, validated six-section H3 video brief that feeds ComfyUI's native H3 render nodes. Three nodes (Main, Media, Setup), a creativity-level dial, strict brief validation, and exact dialogue/reference-image binding via @-syntax prompts.

▣ Special Stuff

  • keys-heretic-MiniMax-H3 sol-engine + speed upgrades + upscaler finish — Single DGX Spark by drowzeys - One-shot recipe for MiniMax-H3 on a single NVIDIA DGX Spark (GB10, sm_121): Sol-Engine ports, Ultra-Heretic TE, Spectrum forecasting, SageAttention, 0.5 MPix generate + RealESRGAN x2 finish. Includes formal benchmark ladder (1.55× vs dense stock).

  • h3.c (h3-metal) by antirez - Native C/Metal inference engine for Apple Silicon. Prompt-to-video/audio, first/last-frame, and Ref2VA references work end-to-end on M3/M5 Max. Interactive Iris-style session. Not a ComfyUI node — standalone binary.

  • Omni-Rewriter by WayneJin0918 - Open agentic prompt-expansion (PE) harness for image/video generation. Turns everyday intent into validated, model-ready prompts via a bounded AI-agent loop (Analyze → Draft → Validate → Repair → Render). Current video profile is MiniMax-H3; ships a CLI (omni-rewriter expand) + HTTP server (POST /v1/expand), deterministic PE validation, and a reusable CI lint Action. Apache-2.0. Not a ComfyUI node — standalone tool (generation adapters stay outside expand).

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Guides & Tutorials

▣ Official Guides

  • Video Prompt Writing Guide (Base) - Official MiniMax-H3 prompt writing guide for base (FL2VA) mode. Covers prompt structure, camera language, scene composition, and best practices for text-to-video and image-to-video generation.
  • Video Prompt Writing Guide (Reference) - Official MiniMax-H3 prompt writing guide for reference (Ref2VA) mode. Covers multi-modal reference inputs, image/video/audio reference handling, and prompt construction for omni-reference generation.

▣ ComfyUI Tutorials

▣ Performance

  • MiniMax H3 — Performance & Best-Configuration Report - Local-inference performance guide for MiniMax H3 (FL2VA / Ref2VA) across consumer & workstation GPUs, Apple Silicon, and the DGX Spark — distilled from 2 hard-numbered benchmarks and 17 community field reports. Covers a TL;DR config recommendation, hardware-tier tiers, the best speed/quality recipe, and caveats & licensing.
  • MiniMax H3 on an RTX 3060 12GB: what we actually measured - Real-world write-up of running MiniMax-H3 on a 12 GB RTX 3060 — what actually fits, at what resolution and step counts, and the configuration that worked.

◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆◇◆

▓ Workflow & Technical Notes

❖ ComfyUI

Official ComfyUI workflow templates for MiniMax-H3:

❖ OrbitQuant

❖ Abiray

❖ Community Packs

  • joeygambino/MiniMax-H3-Multishot-Workflow - Seamless multi-shot chaining workflow for MiniMax-H3 in ComfyUI — string multiple FL2VA/Ref2VA clips into one continuous sequence with matched audio handoffs. Apache-2.0.
  • javawock7618/comfy-MiniMax-H3-workflows - Curated ComfyUI workflow pack covering the full low-VRAM acceleration stack in one importable bundle: INT8 + SageAttention + Spectrum + Lightx2v + Turbo + Motion Context + Latent Upscale + TTS.