Qwen-Image-Edit guide: instruction prompts, multi-image editing and the LoRAs that make it better
Topic: Prompting
By Captain
Published 2026-05-26
How to get clean edits from Qwen-Image-Edit: writing instructions instead of descriptions, keeping identity and layout, multi-image inputs, text replacement, pose and depth control, the lightning LoRAs for 4-step edits and the anything-to-real and slider LoRAs trending in the Qwen2 family.
Overview
Qwen-Image-Edit is the editing model of Alibaba's
Qwen-Image
line: you give it a picture and an instruction, and it returns the picture with that change and nothing else changed. The 2509 release added several input images (put the jacket from image two on the person in image one), better identity preservation, pose and depth guidance and the same text rendering that made Qwen-Image famous, so you can replace a sign or a label with exact words.
In this catalog the editing generation and its
LoRAs
live in the Qwen2 family. Its LoRAs are among the fastest growing on the site: anything-to-real conversions that turn drawings and renders into photographs, age and body sliders, deblurring, watermark removal, consistency LoRAs that keep a character across edits, and the lightning LoRAs that cut an edit from 40 steps to 4.
The model is 20B parameters, so most people run it quantised (fp8 or GGUF) on 12 to 24 GB or hosted:
PirateDiffusion
exposes it as the edit-qwen workflow (/wf /run:edit-qwen with an input image) and
BitVector
has it in the web app with the LoRAs preinstalled.
Reference
Name | Type | What it is |
|---|---|---|
Instruction prompt | text | An imperative sentence: "Change the shirt to red", "Remove the person on the left", "Replace the sign text with OPEN 24H". Not a scene description. |
Image 1 (and 2, 3) | input | The picture to edit; extra images provide a reference object, a style or a second person. Say which image is which in the prompt. |
Steps / CFG | setting | |
Denoise | setting | Keep 1.0; the model decides what to change from the instruction. Lower values blur edits. |
Resolution | setting | The model works near 1 megapixel; the ComfyUI template scales the input and back. Very large inputs are downscaled. |
Pose / depth control | input | A pose skeleton or depth map as an extra image with "follow the pose of image 2". |
Text replacement | prompt | Put the exact text in quotes and name where it goes: Replace the text on the mug with "World's OK Dad". |
Lightning LoRA | LoRA | Qwen-Image-Edit-Lightning 4-step or 8-step; weight 1.0; CFG 1. |
Style and conversion LoRAs | LoRA | anything2real, anime2real, cinematic, morereal: weight 0.8 to 1.0, added to the instruction with their trigger words. |
AD

No install required - run AI models on the cloud
BitVector Prism is the easiest way to start: pick a model, type a prompt and generate in a clean web app, with nothing to configure. BitVector is also available on Discord and on the web (SpyGlass).
Step by step
- Load the ComfyUI Qwen-Image-Edit template (Browse Templates > Image > Qwen Image Edit) and set the loaders: the edit diffusion model (fp8 or GGUF), the Qwen2.5-VL text encoder, theQwenVAE.
- Drop the picture into the Load Image node. Write a single instruction. Keep everything you do not mention unmentioned; the model preserves it.
- Run at 40 steps / CFG 4 once to see the baseline; then add the lightning LoRA from the Qwen2 family page and switch to 8 steps / CFG 1.
- For multi-image edits, add a second Load Image and refer to it: "Put the sunglasses from image 2 on the man in image 1." Reference images work best cropped to the object.
- For text: write the exact string in quotes, name the surface, and ask for the same font or colour if you want continuity.
- For conversions, add the LoRA and its trigger word to the instruction: "anything2real, turn this illustration into a realistic photograph, keep the composition."
- Chain edits: feed the output back in as image 1 for the next instruction instead of asking for three changes at once.
- Hosted: on PirateDiffusion reply to a photo with /wf /run:edit-qwen change the shirt to red, or use /workflow /show:edit-qwen to read the fields; on BitVector pick the edit model and upload.
Examples
Single edits
Change the man's jacket to a dark green wool coat, keep everything else the same
Remove the car in the background and extend the hedge
Replace the text on the storefront sign with "CAFE LUNA" in the same style
Two images
Put the handbag from image 2 on the woman's shoulder in image 1, match the lighting of image 1
Conversion LoRA
anything2real, convert this anime illustration into a realistic photograph of the same scene, same pose and framing, natural skin, 50mm lens
PirateDiffusion
/wf /run:edit-qwen /initimage:Ixyz123 change the shirt to red
Tips
- One change per instruction gives the cleanest result; stack edits in a chain.
- Say what to keep when it matters: "keep her face and hair exactly the same" measurably improves identity.
- Lighting edits ("make it golden hour") work; perspective changes ("show it from above") work in 2509 but need a few seeds.
- Negative promptsare weak here; phrase the instruction positively.
- For product shots, a plain reference of the product as image 2 beats describing it.
- The sliders in the Qwen2 family (age, body, detail) take negative weights for the opposite direction; the model page notes the range.
- Watermark and deblur LoRAs are run with the neutral instruction the author gives ("remove the watermark", "deblur this photo") and nothing else.
Troubleshooting
The whole image changed
Why it happens
A descriptive prompt instead of an instruction, or denoise below 1 on a different workflow.
How to fix it
Write an imperative single change; use the official template with denoise 1.0.
Face drifted
Why it happens
Large edit near the face or a style LoRA at full strength.
How to fix it
Add "keep the face identical", lower the LoRA to 0.7, use a consistency LoRA.
Text has typos
Why it happens
Long string or unusual font request.
How to fix it
Shorter text in quotes, one line, ask for a plain font; run two or three seeds.
Out of memory
Why it happens
20B model in fp16.
How to fix it
fp8 or GGUF Q4 weights and the fp8 text encoder; or hosted.
Edit ignored with the lightning LoRA
Why it happens
CFG left at 4 or steps too low for the LoRA version.
How to fix it
CFG 1, steps matching the LoRA (4 or 8).
Second image not used
Why it happens
The prompt did not reference image 2 explicitly.
How to fix it
Name the images in the instruction.
AD

No install required - run AI models on the cloud
PirateDiffusion is Telegram only and built for pros: thousands of models, LoRAs and workflows driven by chat commands, with unlimited generation on a fixed price plan.
Questions
Qwen-Image-Edit or FLUX.2 for edits?
Qwen for exact text and instruction-following on a single picture;
FLUX.2
for multi-reference compositions and when you also generate from scratch. Both are on this site with Run boxes.
Is it commercial-friendly?
Qwen-Image models are Apache 2.0. Check individual LoRA pages for their licence line.
Why is the family called Qwen2 here?
The catalog groups the editing generation (2509) and its LoRAs apart from the first Qwen-Image release so that LoRAs are never mixed across versions.
Can it edit video frames?
Where do I try it fastest?
PirateDiffusion (/wf /run:edit-qwen in Telegram) or BitVector in the browser; both keep the lightning LoRA installed.
Links and sources
- Qwen-Image on GitHub
- Qwen-Image-Edit-2509 on Hugging Face
- Qwen2 (edit) family page
- Qwen family page
- Hosted edit workflow on PirateDiffusion
- BitVector
Models in this guide
Written by
Captain
Related guides
Prompting
FLUX.2 prompting guide: natural-language prompts, multi-reference, editing and LoRAs
Prompting FLUX.2 dev for the results it is known for: paragraph prompts with subject, scene, light and lens, exact text in quotes, up to several reference images for products, people and styles, instruction-style edits, the guidance and step settings that matter, and how FLUX.2 LoRAs differ from Flux dev LoRAs.
By Captain
2026-05-28
Models
Flux vs SDXL vs SD 1.5 vs Qwen vs Z-Image: which model family to use in 2026
A practical comparison of the open image model families: prompt following, realism, anime and illustration, text rendering, speed, VRAM, LoRA ecosystem and licence, with a recommendation for each kind of work and hardware.
By Captain
2026-05-22
Errors & fixes
CUDA out of memory: how much VRAM each AI model needs and how to fit it
A VRAM table for SD 1.5, SDXL, Flux, FLUX.2, Qwen-Image, Z-Image, Wan 2.2 and HunyuanVideo, and the techniques that make a model fit: fp8 and GGUF quantisation, CPU offload, tiled VAE, resolution and frame limits, and when to stop fighting and run it hosted.
By Captain
2026-05-18
Cloud services
PirateDiffusion: run any model from Telegram, no GPU needed
The commands that matter in the PirateDiffusion Telegram bot: /render with model trigger words, (( )) and [[ ]] weighting, #recipes, /adetailer, /highdef and /facelift upscaling, /remix, /inpaint, ComfyUI workflows with /wf /run:, and the new // skills that pick the model for you.
By Captain
2026-10-03
More guides
Models
FLUX.1 Dev: how to prompt it, best settings and creative ideas
A practical guide to FLUX.1 Dev by Black Forest Labs: natural-language prompting, guidance and step settings, LoRA stacking, text rendering and prompt ideas that play to its strengths.
By Captain
2026-09-30
Models
Stable Diffusion XL 1.0: prompts, settings and what it still does best
How to get the most out of the official SDXL 1.0 base model: native resolutions, CFG and sampler settings, the refiner, negative prompts and ideas for styles where SDXL still shines.
By Captain
2026-10-01
Models
Stable Diffusion 1.5: the classic model, prompted properly
Stable Diffusion 1.5 is still worth knowing: the right resolution, CFG and sampler, how to use its huge library of LoRAs and embeddings, and creative prompt ideas suited to a 512-pixel model.
By Captain
2026-10-02
Models
FLUX.2 Dev: prompting the newest Black Forest Labs model
FLUX.2 Dev brings larger prompts, better text, stronger realism and multi-reference editing. This guide covers the turbo workflow, guidance and resolution settings, prompt structure and ideas that use its new strengths.
By Captain
2026-10-03
Models
Qwen-Image: long prompts, perfect text and bilingual posters
Qwen-Image is the model to reach for when the words in the picture matter. Learn how to write its long descriptive prompts, render English and Chinese text accurately, set CFG and steps, and explore layout-heavy creative ideas.
By Captain
2026-09-29
Models
SDXL-Lightning: four-step generation on any SDXL checkpoint
ByteDance's SDXL-Lightning LoRA turns a 30-step SDXL render into a 4-step one. Learn the step counts, the CFG you must use, the sampler, and how to combine it with your favorite checkpoints and style LoRAs.
By Captain
2026-09-30

Model
Trends
.ai
© ModelTrends.ai
|
Made in Japan
|
© 2026