Downloads · 30 days
0
UCLA-AGI/SPIN-Diffusion-iter1
SPIN-Diffusion-iter1 is a text-to-image model from UCLA-AGI. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as apache-2.0.
Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation (https://huggingface.co/papers/2402.10210)
Downloads · 30 days
0
Access
Public
Updated Feb 24, 2024
Parameters
860M
3.4 GB on disk
Likes
2
Public
Click a slice to open those files.
.safetensors3.4 GB · 100%
From the Hugging Face model README
Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation (https://huggingface.co/papers/2402.10210)

This model is a self-play fine-tuned diffusion model at iteration 1 from runwayml/stable-diffusion-v1-5 using synthetic data based on the winner images of the yuvalkirstain/pickapic_v2 dataset. We have also made a Gradio Demo at UCLA-AGI/SPIN-Diffusion-demo-v1.
The following hyperparameters were used during training:
@misc{yuan2024self,
title={Self-Play Fine-Tuning of Diffusion Models for Text-to-Image Generation},
author={Yuan, Huizhuo and Chen, Zixiang and Ji, Kaixuan and Gu, Quanquan},
year={2024},
eprint={2402.10210},
archivePrefix={arXiv},
primaryClass={cs.LG}
}