Downloads · 30 days
6
46% of all-time downloads
camenduru/IMAGDressing
IMAGDressing is a text-to-image model from camenduru. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as apache-2.0.
Downloads · 30 days
6
46% of all-time downloads
All-time downloads
13
Public
Parameters
860M
18 GB on disk
Likes
1
Public
Click a slice to open those files.
.safetensors9 GB · 50%
From the Hugging Face model README
Project Page | Paper | Code| Data
</div>To address the need for flexible and controllable customizations in virtual try-on systems, we propose IMAGDressing-v1. Specifically, we introduce a garment UNet that captures semantic features from CLIP and texture features from VAE. Our hybrid attention module includes a frozen self-attention and a trainable cross-attention, integrating these features into a frozen denoising UNet to ensure user-controlled editing. We will release a comprehensive dataset, IGv1, with over 200,000 pairs of clothing and dressed images, and establish a standard data assembly pipeline. Furthermore, IMAGDressing-v1 can be combined with extensions like ControlNet, IP-Adapter, T2I-Adapter, and AnimateDiff to enhance diversity and controllability.
