Downloads · 30 days
0
gwang-kim/DiffusionCLIP-CelebA_HQ
DiffusionCLIP-CelebA_HQ is a image-to-image model from gwang-kim. Use it when you need one image transformed into another. It is set up for pytorch.
Creators: Gwanghyun Kim, Taesung Kwon, Jong Chul Ye Paper: https://arxiv.org/abs/2110.02711
Downloads · 30 days
0
Access
Public
Updated Sep 22, 2022
Repo size
455 MB
Likes
11
Public
Click a slice to open those files.
.ckpt455 MB · 100%
From the Hugging Face model README
Creators: Gwanghyun Kim, Taesung Kwon, Jong Chul Ye Paper: https://arxiv.org/abs/2110.02711
<img src="https://github.com/submission10095/DiffusionCLIP_temp/raw/master/imgs/main1.png" alt="Excerpt from DiffusionCLIP paper showcasing comparison of DiffusionCLIP versus other methods for image reconstruction, manipulation, and style transfer." style="height: 300px;"/>DiffusionCLIP is a diffusion model which is well suited for image manipulation thanks to its nearly perfect inversion capability, which is an important advantage over GAN-based models. This checkpoint was trained on the CelebA-HQ Dataset, available on the Hugging Face Hub: https://huggingface.co/datasets/huggan/CelebA-HQ.
This checkpoint is most appropriate for manipulation, reconstruction, and style transfer on images of human faces using the DiffusionCLIP model. To use ID loss for preserving Human face identity, you are required to download the pretrained IR-SE50 model from TreB1eN. Additional information is available on the GitHub repository.
@article{kim2021diffusionclip,
title={Diffusionclip: Text-guided image manipulation using diffusion models},
author={Kim, Gwanghyun and Ye, Jong Chul},
journal={arXiv preprint arXiv:2110.02711},
year={2021}
}