Qwen_VL_8B for Qwen Image 2.1 (QP)
Note: I have tested uncensored and abliltered models on QWEN 2.1 those models tend to misalign and cause aberrations with the 2.1 DiT model. QP does not effect the core LLM and leaves the BF16 intact while predicting FP32, this allows for less rounding errors when quantizing.
All models other then FP16 use Comfy Metadata for quantization.
Requirements:
Update CUDA to 13.4.2
Update Pytorch to 13.2
Update at minimum comfy-aimdo, comfy-kitchen
pytorch version: 2.13.0+cu132
xformers version: 0.0.35
Using xformers attention
ComfyUI version: 0.35.0
comfy-aimdo version: 0.5.5
comfy-kitchen version: 0.2.34
QP (Quantization Prediction)
QP is theoretically improving 40-50% of the blocks on the trailing 16 Matnitsa bits.
For the other 50-60% that it does not improve it did not degrade them more then what they would have been rounded to in the first place.
This was tested on Full FP32 trainings such as T5.
Description
FAQ
Comments (4)
Does this new model already support NSFW content? (Uncensored)
To the extent china allows, it seems to be.


