Downloads · 30 days
1.1K
5% of all-time downloads
joyfox/Qwen-Image-Edit-MeiTu
Qwen-Image-Edit-MeiTu is a image-to-image model from joyfox. Use it when you need one image transformed into another. It is set up for diffusers. The card lists the license as apache-2.0.
<p align="center" <img src="https://ai.static.ad2.cc/banner.png" width="1000"/ </p
Downloads · 30 days
1.1K
5% of all-time downloads
All-time downloads
19.6K
Public
Repo size
84.6 GB
Likes
154
Public
Click a slice to open those files.
.gguf64.1 GB · 76%
From the Hugging Face model README
This model — Qwen-Image-Edit-MeiTu — is an improved variant of Qwen/Qwen-Image-Edit, built with DiT-based architecture fine-tuning to enhance visual consistency, aesthetic quality, and structural alignment in complex edits.
Developed by Valiant Cat AI Lab, this version aims to further close the gap between high-fidelity semantic editing and coherent artistic rendering, achieving a more natural and professional output across a wide range of prompts and subjects.
Enhanced Consistency:
Utilizes DiT (Diffusion Transformer) fine-tuning to ensure structural stability between input and edited regions, maintaining global spatial coherence.
Aesthetic Optimization:
Trained with aesthetic discriminators and curated aesthetic score datasets, producing more pleasing colors, contrast, and light balance.
Better Detail Preservation:
Improved low-level reconstruction for fine details such as textures, faces, and typography.
Broader Scene Adaptability:
Performs well on portraits, environments, product photos, and illustrations, supporting both semantic and appearance-based editing.
Below are examples of consistency and aesthetic improvement in complex editing scenarios:
| Input & Output |
|---|
| <img src="preview/result1.png" width="800"/> |
| <img src="preview/result2.png" width="800"/> |
| <img src="preview/result3.png" width="800"/> |
| <img src="preview/result4.png" width="800"/> |
| <img src="preview/result5.png" width="800"/> |
Try these prompts to explore the model’s strengths:
This model works seamlessly with a modified ComfyUI Qwen-Image-Edit workflow.
Just use this model in the Unet node to workflow for edit image.
Weights available in Safetensors format:
👉 Download Qwen-Image-Edit-MeiTu
This model was trained and optimized by the
AI Laboratory of Chongqing Valiant Cat Technology Co., LTD.
Visit https://vvicat.com/ for business collaborations or research partnerships.
This model is part of the Qwen-Edit+ research line and is associated with the following preprint:
Fan Tang, Siyuan Li
Qwen-Edit+: Scaling Image Editing with VLM-Guided Consistency and Aesthetic Preference Distillation.
Research Square, Version 1, 08 April 2026.
DOI: 10.21203/rs.3.rs-9352857/v1
If you use this model, please cite:
@article{tang2026qweneditplus,
author = {Fan Tang and Siyuan Li},
title = {Qwen-Edit+: Scaling Image Editing with VLM-Guided Consistency and Aesthetic Preference Distillation},
journal = {Research Square},
year = {2026},
doi = {10.21203/rs.3.rs-9352857/v1},
url = {https://doi.org/10.21203/rs.3.rs-9352857/v1}
}
Licensed under Apache 2.0.
We are hiring research engineers and creative ML practitioners at
Chongqing Valiant Cat Technology Co., LTD — reach out via
📧 [email protected]