The final form (probably) of my favorite blend of Krea2 Turbo and the best of the unlocking bypasses. This is (to the best of my knowledge, anyway) as close to the original, not-yet-censored version of Krea2T as it gets, without adding any extra training data or forcing it to hallucinate a roomful of unasked-for details. I did it solely to "improve prompt adherence".
This is a merge of the standard Krea2 Turbo diffusion model, the Fedor bypass, and the TextFusion Refusal-Reduction LoRA. It is available in 5 different sizes/quants, including BF16, in case you want to use it to make your own merges, or even if you just like using OG-sized models and have the VRAM for it (I do not, lol).
Please note: this was made using ComfyUI, and works in ComfyUI, but other inference setups may have issues with some quants. I'm working on figuring out how to make them accessible to more than just ComfyUI users, but I can't promise anything. I'm doing all of this on my own setup at home, and it's not exactly SOTA.
Sometimes an ordinary text encoder will make this model produce water out of nowhere or do other weird things, like replace your realistic subject with a cartoon, make them wear flesh-colored underwear, or cover their skin with a patchwork of weird stickers. If you need an unlocked text encoder, try looking here. It can help a lot. You might still get a cartoon sometimes when prompting for a photograph, or get a weird skin effect, but it happens a lot less frequently.
You'll want a decent VAE too (for image quality); I recommend the FP32 version of the Wan2.1 VAE over the standard Qwen Image VAE. Yes, it's compatible, and it makes nicer hair and fabrics and stuff.
This is my personal favorite Krea2 merge. It knows to leave eyeglasses on nudes. You can prompt it for someone to be nude, then tell it they're wearing something, and they will be nude except for that specified something. It has native knowledge of certain specialty clothing items and things that other models have always required LoRAs for. More often than not you'll get normal-looking hands and feet, even. It's like an evolved SDXL with 10x-better training, minus the brain damage. Not trying to throw shade, but I haven't seen extra legs in over 1000 generations, just saying. It's still not perfect, but then what else is? It still makes nicer bewbies than any other potato-friendly model without any help from a NSFW LoRA.
Sample images were generated using nothing but this base diffusion model (Int8 for most) and simple text prompts. Some used only 4 steps, I think. And yes, I threw some away, because I suck at this stuff. No additional LoRAs were used in making the samples, just my own pathetic prompting abilities. Also, this is just a decensored base model, it doesn't have any added training for aesthetic improvement. This is two-edged; the base model still struggles with some things (like skin details), but you can also use LoRAs without them fighting against the base model's bias, like they would with a merge that already included style LoRAs, detailers, someone else's preferred body types, etc.
Naturally, your mileage may vary with different quants on different machines. I find the Int8 runs faster on my machine than the NVFP4, even though it's larger. Int4 takes me longer to initialize than Int8, but then it gives me ~1.1 seconds per step (which is pretty fast in my world). Google which quant runs best on your computer, and maybe you can save yourself a lot of downloading and testing. And before anybody asks me, no, I can't do an MXFP8 quant. My machine just refuses to do it; believe me, I've tried over 9,000 times. >.<;
And bear in mind, the smaller the quant, the greater the difference in output quality, just as a general rule. This doesn't mean smaller quants are all crap, it just means you might need to use more steps to get better images, use a realism enhancing LoRA, and/or maybe do an upscale/detail fix. Use clear, simple, specific prompts. Where there's a will there's a way, and don't let my shoddy sample images fool you into thinking otherwise.
I've lately been exploring improvements for "photography" with Krea2, and have found that I can actually get pretty realistic outputs even with the Int4 quant. One thing that helps greatly, after much searching for realism enhancers for Krea2, is the Pornmaster Realism Slider. I get very good results with it at a strength of 5.0, and it helps a lot with that random-yet-all-the-same "AI face" thing (and probably more). To improve skin detail (or the lack thereof) Loraholic's Skin Detail Slider can help a lot. When using both sliders, I set Loraholic's to 2.0, and Pornmaster's to 2.5. I also like to do a detailing "upscale"pass with the 1x-ITF-SkinDiffDetail-Lite model or similar (there are some more you can check out here), to help finish off the last of the plasticky skin texture and to help bring out finer hair detail. I find that 8 steps can be way better than 12, though 9 might actually be the sweet spot, just like with some other "turbo" models. It can poop out cartoons and Pixar-style stuff and conceptual quickies at lightning speed with only 4 steps. Also lower-res photographic-style stuff.
Krea2 is licensed under this license, and neither the people who made the model, the people who made the unlocks, nor I are in any way responsible for how you use this model. Please prompt responsibly, because even my crappy sample images can get flagged as indistinguishable from real photos sometimes, and our AI generations can have real-world consequences, regardless of how fun it is to fool artists on Reddit, etc. Be good, and have fun! I know you can do both~! XD
Description
Int8, Int4, NVFP4, FP8, & BF16 available. Look under the download tab for your desired variant. Made with/quantized for use with ComfyUI, so YMMV.
FAQ
Comments (26)
could you add int4 convrot version plwase? >.<
If I can figure out how to make my ComfyUI do it, then yes, I will. I can't make any guarantees, because my setup doesn't let me do everything (practically running a potato myself), but if I can make it happen, I will.
Okay, I think I've managed to pull off an Int4-convRot quantization, but I've never used Int4 for anything and ComfyUI is giving me errors when I try to use it, because I haven't the slightest idea of how to make a workflow for it. I therefore have no idea if it actually works or not, so I would feel bad if I uploaded it and it turned out to be a cold turd.
I will work on this tonight and see if I can find a way to get a workflow together and test it properly. My computer let me make the quant, so I should be able to make it work if it is actually usable. Failing that, I'll upload it and ask if maybe you and others could test it for me and see if it works. Thanks for your patience. ^-^;
To think int 8 is already a little iffy about quality. Int 4 must be the lobotomy
Okay, uploading now. I got it to work, and it's actually Int4 now.
Decent - really mucks up skin details at times (freckles can be splotchy and gross). Does OK with Character LoRA. Thanks for sharing.
Yeah, it's far from perfect, but if I can make it better I'll update it. For some things LoRAs are still way better, and it never hurts to use a good photography/detail LoRA either way.
@PheebyKatz Totally! Thanks for putting together a base that eliminates SOME of the LoRA stack ;)
Works really well for me. But would appreciate a bit more vram friendly version. Any chance of getting a nvfp4 version?
Uploading it now. I can't make any guarantees as to the speed or quality, but it didn't break anything when I ran it. I'm not on a Blackwell, so I get no speed advantage with NVFP4 (it runs like FP8 for me) but it's smaller. My gens are taking about 3-ish seconds per iteration, but if you're running a machine that likes NVFP4, you'll probably have better results than me. Please test it and let me know, and I'll do what I can to make it better, if possible.
It is done. Check the downloads tab, it's there for you.
@PheebyKatz mind me asking what gpu you're using?
@Spoopy123 I'm on an RTX 3060 ti, I have 8GB of VRAM and 16GB of system RAM. Basically mid-grade. I call it my sweet potato, because while it can't handle some larger models it does pretty well regardless. It just sucks that some smaller quants don't run as fast on it as some larger ones do. SDD space is a thing.
Thank you very much. I'm using RTX 5080, so 16GB VRAM. I'll upload a few images and a review, once I've had time to test it.
@AnteMaxx Cool, I hope it works well enough to be worth downloading, at least! ^-^;
Unfortunately I cannot get it to work. I don't use Comfy, but SD Forge Neo, which supports Krea 2 and all other Krea 2 models, including nvfp4 do work on it, but this one doesn't. Tried asking ChatGPT and apparently it isn't something I can fix. "The NVFP4 Krea 2 checkpoint is packed differently than Forge Neo expects, causing incompatible tensor sizes when the model loads."
@AnteMaxx Aw, that stinks. I use the Easy-Install version of ComfyUI because it's given me the least issues out of everything else; too bad it doesn't output stuff that's compatible with everything. It worked for me right out of the quantizer workflow.
At least you can still do your own merge on the fly. Get the bypass LoRAs and wire them into your inference setup, and it'll probably work. I know you can find a Krea2 Turbo model in a quant that will actually work for you, look on Huggingface.
In the meantime, I'll look into whether I can make one that's compatible with more than ComfyUI.
@AnteMaxx you have to use an nvfp4 model on a 16gb gpu? do you run into any issues on forge neo with running int8 or fp8 models (or any ~12gb krea2 model)?
@Spoopy123 When using a nvfp4 model, I don't have to offload anything, and I don't see a noticeable loss in quality. So I try to use them as I can. I am still new to Krea 2, the previous models I used, about 2 years ago, were SDXL models. So my setup may not be optimized in the best possible way.
@AnteMaxx if you haven't tried the latest ComfyUI, maybe give it a go, it can be a real game-changer. I highly recommend it if your system can run it. I'm using the Easy Install version, and since a few updates ago, I'm able to run models that didn't previously fit in my VRAM, and it doesn't take forever for them to load anymore.
My VRAM is only 8GB. I'm using the 12GB version of Krea, about 10 gigs of text encoder, and even with LoRAs tacked on I'm getting fast generation speeds, and it doesn't use the paging file anymore (which made offloading slow), it just uses a placeholder for each model in RAM.
It's not like it was even six months ago, they've really improved it and it's allowed me to use models I'd only wished I could use before. If at all possible, you should try it.
https://blog.comfy.org/p/dynamic-vram-in-comfyui-saving-local <-- Explains it better than I can.
anyone else tried the nvfp4 version yet? I'm getting a vae tuple out of range error... tried multiple different vae. my other nvfp4 models are working fine.
I just tested it again, and it's still working, just slower for me because I'm not on a Blackwell... Are you using Forge Neo? Someone else said they were using it and couldn't get the model to work. It was made in ComfyUI, and works in ComfyUI, at least.
I'll reupload it in case the file got screwed up somehow, but really that's all I can do. I only made the quants besides Int8 in case they were useful to someone else, doing all of this as a hobby in my spare time on a retired gaming desktop.
I wish I could make it work for everybody. It's a great model, and everyone should be able to use it. This is kind of a bummer.
I'm on ComfyUI. I guess it's possible the upload got corrupted, if it's working on your end but not for me and this other user?
it's ok, worse case scenario I try the int8 or fp8. but if you wouldn't mind trying the upload again I'd be appreciative. :)
Okay, I've re-uploaded the file, if you care to try it again. If it won't go, you can still use the unlocks listed in the second paragraph of the main post. Just stick them in a LoRA loader. I use the Fedor bypass at 2.0 and the TextFusion bypass at 1.0, and it seems to give the best results.
Apologies for any inconvenience, or for any file incompatibility issues with your inference setup.
@dow185 Just saw this, I must have been lost in reuploading land, lol. Anyway, it's been re-upped, see if it works. I hope it does.
@PheebyKatz thanks, I got it working now. seems like the issue was trying to load it through 'load checkpoint', and using 'load diffusion model' has it working. which is interesting, since I've been loading every other model through that node without issue - int8/fp8/fp16/nvfp4.

