CivArchive
    Minimax-h3_Singularity - Pruned_int8
    NSFW
    Preview 141966940

    🌐 Online Interactive Demo

    Test the model directly in your browser without local GPU setup:

    👉 Try it on RunningHub Workflows (https://www.runninghub.ai/post/2096339589492432897/?inviteCode=rh-v1559)

    📖 Model Overview

    Minimax-h3_Singularity is a comprehensive fine-tuned fusion model specialized in enhancing the capabilities of MiniMax-H3. Designed as a versatile multimodal video generation model, it natively supports Text-to-Video (T2V), Image-to-Video (I2V), Reference-to-Video (Ref2V), and Video-to-Video (V2V) workflows within ComfyUI.

    Built upon a strategic fusion of key checkpoints (including ref, fl, b25-49, etc.), this model underwent deep high-step fine-tuning. To preserve the original model's foundational strengths and broad generalization while solving artifacts introduced by high-step training, we spent 3 full days on precise model pruning and weight optimization. The result is a clean, sharp, and highly dynamic video generation model.

    Key Improvements & Features

    🎬 HDR Image Quality & Blur Reduction: Fine-tuned on high-dynamic-range (HDR) video datasets to significantly enhance visual clarity and eliminate motion blur during high-speed action.

    👤 Distant Face Restoration: Drastically reduces facial distortion, blurriness, and collapsing in medium-to-long shots.

    🎨 Clean & De-Oiled Aesthetic: Removes heavy, unnatural skin shine and glossy textures, rendering natural lighting and photorealistic materials.

    ⚔️ Enhanced Dynamic Motion: Boosts motion fluidity and physical impact, excels in complex action sequences such as sword fighting and martial arts/melee combat.

    🌌 VFX & Fantasy Effects: Specifically optimized for fantasy spellcasting, particle aura, and magical combat visual effects.

    🎭 Expressive Facial Dynamics: Captures subtle facial expressions and emotional nuances more vividly.

    📹 Cinematography & Camera Control: Strengthens responsiveness to camera movements (pan, tilt, zoom, tracking shots) for cinematic storytelling.

    🛡️ Full Base Capability Retention: 100% preserves MiniMax-H3's original prompt adherence, style adaptability, and base multimodal generation strength.

    💡 Usage Guide

    Multimodal Pipeline Support

    This model is fully compatible with ComfyUI and supports:

    • Text-to-Video (T2V)

    • Image-to-Video (I2V)

    • Reference-to-Video (Ref2V)

    • Video-to-Video (V2V)

    • 🚀 Recommended Acceleration LoRA

    For high-speed generation with minimal quality loss, we strongly recommend pairing with:

    • minimax_h3_ref2v_turbo_4step_v0.1 (Enables 4-step fast inference)

    🙏 Acknowledgements

    Special thanks to the MiniMax open-source team for creating and releasing the powerful MiniMax-H3 multimodal video model, providing a solid foundation for the open-source community! 🤝

    🤝 Community & Commercial Inquiries

    Feel free to connect for tutorials, community discussions, workflow sharing, or commercial collaborations:

    • YouTube Channel: AIGC-Singularity (https://youtube.com/@AIGC-Singularity)

    • Bilibili Channel: AIGC-Singularity Space

    • QQ Group 1: 1058747239 (Request to join)

    • QQ Group 2: 1072010342 (Request to join)

    • Business Inquiries (WeChat): aigctyd

    • Email: [email protected]

    Description

    FAQ

    Comments (4)

    FourBunnySep 6, 2026
    CivitAI

    playnx09699Sep 6, 2026
    CivitAI

    Although I haven’t had much chance to test it yet, there are many areas where this model has improved to such an extent that it’s impossible to compare it with other models. In particular, its emotional expression and dialogue handling have become outstanding, and the sound quality has improved significantly. (It also seems to handle NSFW content exceptionally well; I’d recommend setting aside all other LORA models and giving this one a try.)

    takomli2013979Sep 6, 2026
    CivitAI

    Tested a bit, the improvement is significant. with Other model, it have difficulty to tune with 2 d anime style and realistic style appear in same video at same time, and this one can do it well. the audio also seem more lively. for personal feel, it improve the reference consistence in a very good way. btw, why put this in unet section but not checkpoint section? it's not a gguf

    lolmao500Sep 6, 2026· 1 reaction
    CivitAI

    From the comments it looks good. Is it possible to have a FP8 version or maybe a Q6 version so it can run on AMD GPUs with 16gb of vram? Thanks

    UNet
    MiniMax H3

    Details

    Downloads
    176
    Platform
    CivitAI
    Platform Status
    Available
    Created
    9/6/2026
    Updated
    9/8/2026
    Deleted
    -

    Files

    minimaxH3Singularity_prunedInt8.safetensors