Downloads · 30 days
19
1% of all-time downloads
kvablack/ddpo-alignment
ddpo-alignment is a text-to-image model from kvablack. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as creativeml-openrail-m.
This model was finetuned from Stable Diffusion v1-4 using DDPO and a reward function that uses LLaVA to measure prompt-image alignment. See the project website for more details.
Downloads · 30 days
19
1% of all-time downloads
All-time downloads
1.6K
Public
Repo size
8.9 GB
Likes
7
Public
Click a slice to open those files.
.bin5.5 GB · 100%
From the Hugging Face model README
This model was finetuned from Stable Diffusion v1-4 using DDPO and a reward function that uses LLaVA to measure prompt-image alignment. See the project website for more details.
The model was finetuned for 200 iterations with a batch size of 256 samples per iteration. During finetuning, we used prompts of the form: "a(n) <animal> <activity>". We selected the animal and activity from the following lists, so try those for the best results. However, we also observed limited generalization to other prompts.
Activities:
Animals: