Base MinimaxH3 (Turbo Merge)
This model is unchanged/untrained from the base model listed. The model was quantized from the source FP32/BF16 model. BOTH THE FL2VA and REF2VA models are 8 step models.
Requirements:
Update CUDA to 13.4.2
Update Pytorch to 13.2
Update at minimum comfy-aimdo, comfy-kitchen
Basic workflow shows how to use both first frame text guided and first frame last frame with comfy kitchen backend node to speed up generation by 60% (This prevents pyattention fallback which is slow)
pytorch version: 2.13.0+cu132
xformers version: 0.0.35
Using xformers attention
ComfyUI version: 0.35.0
comfy-aimdo version: 0.5.5
comfy-kitchen version: 0.2.34
Consider --disable-dynamic-vram if you are having OOM issues after a few generation or crash when trying to use comfy kitchen vs pyattention
Note: INT Convrot Group size is 64 vs 256 on the comfy supplied models
NF4 version of the reference to video/audio model works but the quality is very low