Wuli-art/Gemma-4-for-Qwen-Image-Edit-2511-Prompt-Extend
<div align="center"> <img src="https://img.alicdn.com/imgextra/i4/O1CN011xTzi6280yXyGmk9b_!!6000000007871-0-tps-3072-938.jpg" width=300 /> </div>
Prompt extend plays a critical role in image editing by helping models better understand editing intent and produce more accurate and stable results.
This model is fine-tuned from gemma-4-12B-it for prompt extend with Qwen Image Edit 2511. It takes an editing prompt together with input images and generates an enhanced editing prompt. The model is trained with Prompt Extend Reinforcement Learning (PERL) using ROLL, with Kimi K2.6 serving as the reward worker to evaluate editing results.
<div align="center"> <img src="https://img.alicdn.com/imgextra/i2/O1CN01EwAJvGc1D4D3OTaP_!!6000000001034-2-tps-1672-941.png" width=600 /> </div>
In real-world scenarios, this model delivers better and more stable editing results across a wide range of complex edit prompts. For now, we recommend using this model together with the Qwen Image Edit 2511 8steps LoRA from lightx2v for end-to-end image editing. Please refer to the provided ComfyUI workflow for more details.
ComfyUI Workflow
A ready-to-use ComfyUI workflow is provided here.
Model Download
For models provided by Comfy-Org, you may use any compatible precision or quantized variant. This guide uses the original BF16 version for demonstration.
Download the required model files and place each one in the corresponding ComfyUI directory listed below.
Once downloaded, your ComfyUI model directory should have the following structure:
ComfyUI/
└── models/
├── text_encoders/
│ ├── gemma4_12b_qwen_image_edit_2511_pe.safetensors
│ └── qwen_2.5_vl_7b.safetensors
├── vae/
│ └── qwen_image_vae.safetensors
├── diffusion_models/
│ └── qwen_image_edit_2511_bf16.safetensors
└── loras/
└── Qwen-Image-Edit-2511-Lightning-8steps-V1.0-bf16.safetensorsUsage
- Download all required model files and place them in the directories shown above.
- Restart ComfyUI or refresh the model list.
- Download the workflow JSON file and drag it onto the ComfyUI canvas.
- Make sure each loader node points to the corresponding downloaded model.
- Load one or more input images.
- Enter the original editing instruction.
- Run the workflow to obtain the extended prompt.
Result Visualization
The table below presents editing results of Qwen Image Edit 2511 based on four different prompt extend methods:
- No prompt extend
- Use Qwen3.7 Plus from Bailian
- Use the original gemma-4-12B-it
- Use the fine-tuned gemma-4-12B-it in this repo
All prompt extend methods use the same system prompt from the official GitHub Repo.
