Brief
Qwen-Image-Edit adds precise text editing to image editing
Qwen-Image-Edit is an image editing version of Qwen-Image, built on the 20B Qwen-Image model. It extends text rendering capabilities to editing and combines semantic and appearance control.
Qwen-Image-Edit is built on the 20B Qwen-Image model and extends its text rendering capabilities to image editing, enabling precise text editing. It feeds the input image into Qwen2.5-VL for visual semantic control and into the VAE Encoder for visual appearance control. This allows both semantic and appearance editing in one tool.
Source details and supporting facts
Each line is stated by the page named above it.
Stated by qwenlm.github.io
- Qwen-Image-Edit is the image editing version of Qwen-Image.
- It is built upon the 20B Qwen-Image model.
- It extends Qwen-Image’s unique text rendering capabilities to image editing tasks, enabling precise text editing.
- It simultaneously feeds the input image into Qwen2.5-VL (for visual semantic control) and the VAE Encoder (for visual appearance control).
- It achieves capabilities in both semantic and appearance editing.
Sources
- Qwen blog (Alibaba)Text stored 15 September 2026
How this story was checked. Written from the 1 page listed above, stored 15 September 2026; claims checked against that stored text on 15 September 2026.
What that means
- 5 of 5 reported statements were confirmed against the page that carries them; the rest were removed rather than published.
- Figures in the text were required to appear in the stored source text: yes. Identifiers: yes.
- The check reads stored text only: no claim rests on a fresh look that did not happen.
- Where the reporting was silent, the text says so instead of filling the gap.