CivArchive
    Qwen_VL_8B for Qwen Image 2.1 (QP) INT8, MXFP8, NVFP4 - v1.0
    NSFW
    Preview 144400819
    Preview 144399452
    Preview 144400182

    Qwen_VL_8B for Qwen Image 2.1 (QP)

    Note: I have tested uncensored and abliltered models on QWEN 2.1 those models tend to misalign and cause aberrations with the 2.1 DiT model. QP does not effect the core LLM and leaves the BF16 intact while predicting FP32, this allows for less rounding errors when quantizing.

    All models other then FP16 use Comfy Metadata for quantization.

    Requirements:

    1. Update CUDA to 13.4.2

    2. Update Pytorch to 13.2

    3. Update at minimum comfy-aimdo, comfy-kitchen

    pytorch version: 2.13.0+cu132

    xformers version: 0.0.35

    Using xformers attention

    ComfyUI version: 0.35.0

    comfy-aimdo version: 0.5.5

    comfy-kitchen version: 0.2.34


    QP (Quantization Prediction)

    QP is theoretically improving 40-50% of the blocks on the trailing 16 Matnitsa bits.

    For the other 50-60% that it does not improve it did not degrade them more then what they would have been rounded to in the first place.

    This was tested on Full FP32 trainings such as T5.

    Description

    FAQ

    Comments (4)

    wyxzddsjj919Oct 2, 2026
    CivitAI

    Does this new model already support NSFW content? (Uncensored)

    Felldude
    Author
    Oct 2, 2026

    To the extent china allows, it seems to be.

    mango22Oct 2, 2026
    CivitAI

    does it work on Mac ?

    Felldude
    Author
    Oct 2, 2026

    fp16 should natively

    TextEncoder
    Qwen 2.1

    Details

    Downloads
    298
    Platform
    CivitAI
    Platform Status
    Available
    Created
    10/1/2026
    Updated
    10/5/2026
    Deleted
    -

    Files

    qwenVL8BForQwenImage21QP_v10_fp16.safetensors

    qwenVL8BForQwenImage21QP_v10_int8.safetensors

    qwenVL8BForQwenImage21QP_v10_mxfp8.safetensors

    qwenVL8BForQwenImage21QP_v10_nvfp4.safetensors