Downloads · 30 days
13
24% of all-time downloads
FlameF0X/NanoDiffusion-46M-1k
NanoDiffusion-46M-1k is a text-to-image model from FlameF0X. Use it when you need an image from a text prompt. The card lists the license as apache-2.0.
A ~46M parameter UNet trained from scratch for text-to-image generation in latent space (SD VAE + CLIP text encoder, both frozen).
Downloads · 30 days
13
24% of all-time downloads
All-time downloads
54
Public
Parameters
45.5M
364 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors182 MB · 100%
From the Hugging Face model README
A ~46M parameter UNet trained from scratch for text-to-image generation in latent space (SD VAE + CLIP text encoder, both frozen).
Dataset: BitTranslate/Bittensor_subnet_19_06_04_24
Trained steps: 1000
Image size: 256px → 32×32 latents
runwayml/stable-diffusion-v1-5 vae subfolderrunwayml/stable-diffusion-v1-5from train_mini_ldm import generate
imgs = generate("a sunset over the ocean", "./mini_ldm_output/final")