CivArchive

    Minimax H3 INT8/INT4 Convrot

    Required Components added to description as the files link to the official Civitai Minimax H3 model page which is set to generation-only, with no option to unlink.
    text_encoders/qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors
    vae/minimax_h3_audio_vae_fp32.safetensors
    vae/minimax_h3_video_vae_fp16.safetensors

    Just uploaded both FL2VA & REF2VA w4a8_mixed models quantized by Kijai. (Makes my mixed int4 models obsolete) True int4 model size with int8 activations, near int8 quality, same speed as int8. For RAM/VRAM-constrained systems. Update your ComfyUI to the latest version for support of the new w4a8 (Should be included in Stable ComfyUI v0.31.0 https://github.com/Comfy-Org/ComfyUI/pull/15308

    FL2VA - first last (frame) to video / audio
    REF2VA - ref_images / ref_videos / ref_video_audios / ref_audios: up to 9 reference images, 3 reference videos (each may carry its own paired soundtrack), and 3 standalone reference audio clips

    Both models can generate t2v (Text to video), i2v (Image to video), v2v (Video to Video), a2v (Audio to video), and multiple references (image/video/audio). But were further fine-tuned/trained for higher-quality outputs for the intended use.

    Description

    True int4 model size with int8 activations, near int8 quality same speed as int8. For ram contsrained systems

    FAQ

    Comments (14)

    herkus_baronas631Aug 8, 2026· 1 reaction
    CivitAI

    For me this type of model dont work at all. Or maybe is very, very slow... i wait near 5 minutes, but still zero movement... GPU tepmerature was low, so its mean dont render anything...

    gambikules858Aug 8, 2026

    how many Vram or Ram.

    Vram 16, Ram 64.

    tsolful
    Author
    Aug 8, 2026

    @herkus_baronas631 Try watching this custom node as your models load, they could be stalling while loading. INT8 and the new Asym_A4W8 works good for me on 3060 12gb 32gb ram. Generation time on the first example of the asym quant model in the prompt. https://github.com/kijai/ComfyUI-MemoryVisualization

    herkus_baronas631Aug 8, 2026· 1 reaction

    @tsolful Ok, would try. TY.

    tsolful
    Author
    Aug 8, 2026· 1 reaction

    @herkus_baronas631 How'd it go, if it is getting stuck while loading and your on a single gpu add --vram-headroom 2 to your startup bat file

    herkus_baronas631Aug 8, 2026· 1 reaction

    @tsolful I haven't tried it, I haven't had time. When I try it, I'll write.

    bhoppingAug 8, 2026· 1 reaction

    I had the same issue. I also have 16GB VRAM and 32GB RAM, and I had to lower both the resolution preset and duration to get it running.

    I'm using the INT8 pruned model with the Speed LoRA at 0.7 strength, 11 steps, the 0.52 MP resolution preset, and 8 seconds. It takes around 5 minutes or more to generate for me.

    Without the Speed LoRA, I was able to get away with 20–30 steps at 0.63 MP, but I had to drop the duration to around 5 seconds, and generations were taking 10+ minutes.

    So if you're getting zero movement for several minutes, I'd try lowering the resolution and duration first and just let it sit for a while. My GPU temperature is around 60°C, with about 94% VRAM usage and around 70% GPU utilization in Task Manager.

    I actually have more trouble trying to run the Ref2V model.

    Speed LoRA: https://civitai.red/models/2837571/minimax-h3-turbo-loras?modelVersionId=3202732

    herkus_baronas631Aug 8, 2026· 1 reaction

    @tsolful Thank you for KJ node for vram and ram. By feelings, i think now work faster. But this new type of quantisation dont work. Ram and vram isnt full, but temperature down, gpu on 99 percent and nothing... so for me INT8 pruned is best solution.

    AwkwardMove67204841Aug 8, 2026
    CivitAI

    Давнго пора было взломать эту штуку надоело платить им

    jynkz41Aug 8, 2026· 1 reaction
    CivitAI

    just as a heads up.. both w4a8 checkpoints you uploaded are actually the ref2va checkpoints

    tsolful
    Author
    Aug 8, 2026

    For me the they are correct, and fl2va is the 11.68gb ref2va is the 10.96gb
    Edit: They are different files, contacted support as they are downloading as the same name

    RJY17Aug 8, 2026
    CivitAI

    Is the latest Asymw4a8 model supposed to be slower than your first int4b_q model ?

    It doubles the time it takes to generate

    xiaobai_solonAug 9, 2026· 1 reaction
    CivitAI

    Many thanks to the developers for creating such an outstanding model and open‑sourcing it. I am still in the process of learning. During practical operations, I find it difficult to control the fine details of camera movement. Could the developers and experts in the community please take a look at this design? Is it feasible to implement this feature as a plugin? I created a somewhat flawed schematic diagram using nano2. https://civitai.red/images/139169154

    Checkpoint
    MiniMax H3

    Details

    Downloads
    916
    Platform
    CivitAI
    Platform Status
    Available
    Created
    8/8/2026
    Updated
    8/13/2026
    Deleted
    -

    Files

    minimaxH3INT8INT4_flREF2VAPruned.safetensors

    minimaxH3INT8INT4_flREF2VAPruned.json

    Mirrors