Int8convrot, converted from https://huggingface.co/circlestone-labs/Anima/blob/main/split_files/diffusion_models/anima-base-v1.0.safetensors
Int8convrot qwen3 0.6b, mainly for fun. Because why not.
qwen3 0.6b (b): 200MiB smaller than previous version. Embeddings also quantized to int8.
Description
FAQ
Comments (6)
I just use this in my workflow but this actually slower than the original model. How is that?
Can it be used with the base model, or only in int8?



