MiniMax H3 pruned fp8 scaled
Pruned FP8 version of the video model
fl2va:
(I2V, FL2VA, T2V)
ref2va:
-Other required files:
nvfp4 Text Encoder: https://huggingface.co/Comfy-Org/MiniMax-H3/resolve/main/text_encoders/qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors?download=true
Description
MiniMax H3 fl2va pruned fp8 scaled
FAQ
Comments (12)
What the difference between int8 and fp8,the latter more fast?
In simplified terms: FP8 is more precise, it uses floating-point. INT8 uses integral.
int8 can be slightly faster depending on hardware.
fp8 in 40 series and up , for others int8 then w4a8 then GGUF if both vram and ram is low
576*928 5秒视频在5060Ti 16G / 32G内存上跑了5:27
I generated one 3s clip with this, and instantly yeeted like 200 GB worth of LTX and Wan models. Holy diff O_O
Yeah. I can't even believe it myself... Like... Uncensored, works, very well, righ tout of the gate. I'm actually finding it unreal to be honest.
@DaddyWolfgang lol, same, and also the fact that you can do so much with basic ahh prompts @.@ Not having to feed it a novel just to control the camera feels like the years of my lifespan stolen by LTX are coming back
Guess I gotta eat crow. I never expected an uncensored model from that company. EVER. Glad I didn't put money on it being censored or not because I'd have lost.
LOL
On a 4070 12GB + 48GB - Generated a 5 seconds 544p clip in just 105 secs with the experimental Turbo LORA, upscaled to 720p (ish) with RTX Super Res. Not as fast as LTX, but the prompt adherance is WAYYYYY better.
@I_XXIV Hi! Please share your workflow.
Details
Files
minimax_h3_video_vae_fp16.safetensors
Mirrors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
minimax_h3_video_vae_fp16.safetensors
