Downloads · 30 days
100
0% of all-time downloads
bguisard/stable-diffusion-nano-2-1
stable-diffusion-nano-2-1 is a text-to-image model from bguisard. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as creativeml-openrail-m.
Stable Diffusion Nano was built during the JAX/Diffusers community sprint 🧨.
Downloads · 30 days
100
0% of all-time downloads
All-time downloads
726K
Public
Parameters
866M
15.5 GB on disk
Likes
19
Public
Click a slice to open those files.
.bin5.2 GB · 50%
From the Hugging Face model README
Stable Diffusion Nano was built during the JAX/Diffusers community sprint 🧨.
Based on stable diffusion and fine-tuned on 128x128 images, Stable Diffusion Nano allows for fast prototyping of diffusion models, enabling quick experimentation with easily available hardware.
It performs reasonably well on several tasks, but it struggles with small details such as faces.
prompt: A watercolor painting of an otter

prompt: Marvel MCU deadpool, red mask, red shirt, red gloves, black shoulders, black elbow pads, black legs, gold buckle, black belt, black mask, white eyes, black boots, fuji low light color 35mm film, downtown Osaka alley at night out of focus in background, neon lights

All parameters were initialized from the stabilityai/stable-diffusion-2-1-base model. The unet was fine tuned as follows:
U-net fine-tuning:
This model is open access and available to all, with a CreativeML OpenRAIL-M license further specifying rights and usage. The CreativeML OpenRAIL License specifies: