Say what to change in plain words, and Qwen Image 2.1 edits your picture without shifting it.
Why use this workflow
You can edit a picture with a plain Qwen Image 2.1 graph and a typed request. This workflow adds three things:
It puts the edit back in place. Qwen draws an edit a little zoomed in or shifted, most of all on style changes. Realign to Source measures that and moves the edit back on top of your original. In my tests, style edits sat 35 px off at the median. After Realign: under 5 px.
It writes the instruction for you. Qwen3-VL looks at your picture and turns a short request into a full instruction. It changes only what you name.
It gives the edit room. A mirrored margin goes around your picture first, so the edit can move without leaving an empty strip at an edge.
Two versions
v1.2 Fast does the edit in 6 steps with the Viggle Turbo LoRA. A run takes about 11 seconds on an RTX 5090. It needs one more file.
v1.1 does it in 25 steps. A run takes about 18 seconds. No extra file.
The edit lines up the same in both.
How to use it
Load your image.
Say the change in plain words, like "make it a soft watercolor painting".
Run.
What you need
AusBoss nodes 2.3.0 or newer. In ComfyUI-Manager, search "AusBoss". Already have them? Update first.
The Qwen Image 2.1 models, on a recent ComfyUI (I tested 0.37.0 and 0.38.0):
qwen_image_2.1_int8_convrot.safetensors →
models/diffusion_models· 7.26 GBqwen3vl_8b_int8_convrot.safetensors →
models/text_encoders· 9.35 GBqwen_image_2.1_vae_bf16.safetensors →
models/vae· 676 MB
For v1.2 Fast: Qwen-Image-2.1-viggle-turbo-v0.3-6step-lora-r128.safetensors →
models/loras· 680 MBOptional: qwen-image-2.1-consistency.safetensors →
models/loras· 159 MB. Its row is off, so it runs without it.
Tips
Say where a moved arm or hand ends up. "put her arm on the table" kept her clasped hands and added a third one. "put her arm on the table, her hands aren't clasped anymore" worked.
Check what Qwen got when an edit misses. Double-click Write the instruction to read the full instruction, then say the missing part more plainly.
The Consistency LoRA holds shapes on restyles. Turn its row on for those, and set Load Image + Pad's four pads to 0 inside Edit with Qwen Image 2.1. Leave it off for pose changes: it holds the old pose.
Realign fixes the frame, not redrawn details. Something the edit redrew in a new spot stays there.
In v1.2 Fast, leave the Viggle Turbo row on. Six steps without it come out soft and washed out.
Credits
Qwen Image 2.1 is by Alibaba's Qwen team, under the Qwen Research License: non-commercial use only. Comfy-Org repackaged the files. The speed LoRA in v1.2 Fast is Viggle Turbo v0.3 by Viggle, under the same license: research use only. The Consistency LoRA is mine, under the same license.
v1.2 Fast (2026-10-04): the same workflow with the edit in 6 steps, about 11 seconds a run instead of 18. v1.1 (2026-09-30): the same edit in a smaller, tidier workflow, and the prompt writer changes only what you ask for. v1.0 (2026-09-28): first release.
Questions or bugs: GitHub issues or the comments here. More from me on X @Zanzibased and GitHub.
Description
The same workflow as v1.1 with one change: the Viggle Turbo LoRA does the edit in 6 steps instead of 25.
A run takes about 11 seconds on an RTX 5090 instead of about 18, with the models loaded.
The edit lines up the same. I ran 8 edits both ways with the same seed.
It needs one more file: Qwen-Image-2.1-viggle-turbo-v0.3-6step-lora-r128.safetensors, 680 MB, in models/loras. The run stops and names the file if it is missing.
v1.1 stays on the page for 25 steps with no extra download.
Needs AusBoss nodes 2.3.0 or newer, as before.
