Minimax H3 INT8/INT4 Convrot
Required Components added to description as the files link to the official Civitai Minimax H3 model page which is set to generation-only, with no option to unlink.
text_encoders/qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors
vae/minimax_h3_audio_vae_fp32.safetensors
vae/minimax_h3_video_vae_fp16.safetensors
Just uploaded both FL2VA & REF2VA w4a8_mixed models quantized by Kijai. (Makes my mixed int4 models obsolete) True int4 model size with int8 activations, near int8 quality, same speed as int8. For RAM/VRAM-constrained systems. Update your ComfyUI to the latest version for support of the new w4a8 (Should be included in Stable ComfyUI v0.31.0 https://github.com/Comfy-Org/ComfyUI/pull/15308
FL2VA - first last (frame) to video / audio
REF2VA - ref_images / ref_videos / ref_video_audios / ref_audios: up to 9 reference images, 3 reference videos (each may carry its own paired soundtrack), and 3 standalone reference audio clips
Both models can generate t2v (Text to video), i2v (Image to video), v2v (Video to Video), a2v (Audio to video), and multiple references (image/video/audio). But were further fine-tuned/trained for higher-quality outputs for the intended use.
Description
True int4 model size with int8 activations, near int8 quality same speed as int8. For ram contsrained systems
FAQ
Comments (14)
For me this type of model dont work at all. Or maybe is very, very slow... i wait near 5 minutes, but still zero movement... GPU tepmerature was low, so its mean dont render anything...
how many Vram or Ram.
Vram 16, Ram 64.
@herkus_baronas631 Try watching this custom node as your models load, they could be stalling while loading. INT8 and the new Asym_A4W8 works good for me on 3060 12gb 32gb ram. Generation time on the first example of the asym quant model in the prompt. https://github.com/kijai/ComfyUI-MemoryVisualization
@tsolful Ok, would try. TY.
@herkus_baronas631 How'd it go, if it is getting stuck while loading and your on a single gpu add --vram-headroom 2 to your startup bat file
@tsolful I haven't tried it, I haven't had time. When I try it, I'll write.
I had the same issue. I also have 16GB VRAM and 32GB RAM, and I had to lower both the resolution preset and duration to get it running.
I'm using the INT8 pruned model with the Speed LoRA at 0.7 strength, 11 steps, the 0.52 MP resolution preset, and 8 seconds. It takes around 5 minutes or more to generate for me.
Without the Speed LoRA, I was able to get away with 20–30 steps at 0.63 MP, but I had to drop the duration to around 5 seconds, and generations were taking 10+ minutes.
So if you're getting zero movement for several minutes, I'd try lowering the resolution and duration first and just let it sit for a while. My GPU temperature is around 60°C, with about 94% VRAM usage and around 70% GPU utilization in Task Manager.
I actually have more trouble trying to run the Ref2V model.
Speed LoRA: https://civitai.red/models/2837571/minimax-h3-turbo-loras?modelVersionId=3202732
@tsolful Thank you for KJ node for vram and ram. By feelings, i think now work faster. But this new type of quantisation dont work. Ram and vram isnt full, but temperature down, gpu on 99 percent and nothing... so for me INT8 pruned is best solution.
Давнго пора было взломать эту штуку надоело платить им
just as a heads up.. both w4a8 checkpoints you uploaded are actually the ref2va checkpoints
For me the they are correct, and fl2va is the 11.68gb ref2va is the 10.96gb
Edit: They are different files, contacted support as they are downloading as the same name
Is the latest Asymw4a8 model supposed to be slower than your first int4b_q model ?
It doubles the time it takes to generate
Many thanks to the developers for creating such an outstanding model and open‑sourcing it. I am still in the process of learning. During practical operations, I find it difficult to control the fine details of camera movement. Could the developers and experts in the community please take a look at this design? Is it feasible to implement this feature as a plugin? I created a somewhat flawed schematic diagram using nano2. https://civitai.red/images/139169154