# Qwen Image 2.1 [MXFP8] 🚀
This is the optimized MXFP8 quantized version of the powerful Qwen Image 2.1 model. The goal of this upload is to bring the excellent visual generation and prompt comprehension capabilities of Qwen 2.1 to setups with limited resources, drastically reducing VRAM consumption with no noticeable loss in quality.
## ✨ Main Highlights
* MXFP8 Efficiency: Utilizes the Microscaling FP8 format to compress the model. This means it takes up almost half the disk space and VRAM compared to FP16/BF16 versions, while maintaining detail precision and structural fidelity.
* Accelerated Inference: Modern GPUs (RTX 5000/4000/3000 series and equivalents) benefit greatly from FP8 compute, resulting in significantly faster generation and processing times.
* Accessibility: Perfect for running locally on graphics cards with lower VRAM, allowing for heavier workflows or higher resolutions that would normally cause an Out of Memory (OOM) error on the base model.
