Downloads · 30 days
80
3% of all-time downloads
Runware/Kandinsky-5.0-T2I-Lite-sft-Diffusers
Kandinsky-5.0-T2I-Lite-sft-Diffusers is a text-to-image model from Runware. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as mit.
<div align="center" <picture <img src="assets/KANDINSKYLOGO1BLACK.png" </picture </div
Downloads · 30 days
80
3% of all-time downloads
All-time downloads
2.3K
Public
Parameters
6B
35.6 GB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors30.5 GB · 86%
From the Hugging Face model README
Kandinsky 5.0 is a family of diffusion models for video and image generation.
Kandinsky 5.0 Image Lite is a lightweight text-to-image (T2I) generation model with 6B parameters.
The model introduces several key innovations:
The original codebase can be found at kandinskylab/Kandinsky-5.
Kandinsky 5.0 Image Lite:
| model_id | Description | Use Cases |
|---|---|---|
| <a href="https://huggingface.co/kandinskylab/Kandinsky-5.0-T2I-Lite-sft-Diffusers">kandinskylab/Kandinsky-5.0-T2I-Lite-sft-Diffusers</a> | 6B supervised fine-tuned text-to-image model | Highest generation quality |
| <a href="https://huggingface.co/kandinskylab/Kandinsky-5.0-I2I-Lite-sft-Diffusers">kandinskylab/Kandinsky-5.0-I2I-Lite-sft-Diffusers</a> | 6B supervised fine-tuned image-to-image editing model | Highest generation quality |
| <a href="https://huggingface.co/kandinskylab/Kandinsky-5.0-T2I-Lite-pretrain-Diffusers">kandinskylab/Kandinsky-5.0-T2I-Lite-pretrain-Diffusers</a> | 6B base pretrained text-to-image model | Research and fine-tuning |
| <a href="https://huggingface.co/kandinskylab/Kandinsky-5.0-I2I-Lite-pretrain-Diffusers">kandinskylab/Kandinsky-5.0-I2I-Lite-pretrain-Diffusers</a> | 6B base pretrained image-to-image editing model | Research and fine-tuning |
import torch
from diffusers import Kandinsky5T2IPipeline
# Load the pipeline
model_id = "kandinskylab/Kandinsky-5.0-T2I-Lite-sft-Diffusers"
pipe = Kandinsky5T2IPipeline.from_pretrained(model_id)
_ = pipe.to(device="cuda", dtype=torch.bfloat16)
# Generate image
prompt = "A fluffy, expressive cat wearing a bright red hat with a soft, slightly textured fabric. The hat should look cozy and well-fitted on the cat’s head. On the front of the hat, add clean, bold white text that reads “SWEET”, clearly visible and neatly centered. Ensure the overall lighting highlights the hat’s color and the cat’s fur details."
output = pipe(
prompt=prompt,
negative_prompt="",
height=1024,
width=1024,
num_inference_steps=50,
guidance_scale=3.5,
).image[0]
@misc{kandinsky2025,
author = {Alexander Belykh and Alexander Varlamov and Alexey Letunovskiy and Anastasia Aliaskina and Anastasia Maltseva and Anastasiia Kargapoltseva and Andrey Shutkin and Anna Averchenkova and Anna Dmitrienko and Bulat Akhmatov and Denis Dimitrov and Denis Koposov and Denis Parkhomenko and Dmitrii and Ilya Vasiliev and Ivan Kirillov and Julia Agafonova and Kirill Chernyshev and Kormilitsyn Semen and Lev Novitskiy and Maria Kovaleva and Mikhail Mamaev and Mikhailov and Nikita Kiselev and Nikita Osterov and Nikolai Gerasimenko and Nikolai Vaulin and Olga Kim and Olga Vdovchenko and Polina Gavrilova and Polina Mikhailova and Tatiana Nikulina and Viacheslav Vasilev and Vladimir Arkhipkin and Vladimir Korviakov and Vladimir Polovnikov and Yury Kolabushin},
title = {Kandinsky 5.0: A family of diffusion models for Video & Image generation},
howpublished = {\url{https://github.com/kandinskylab/Kandinsky-5}},
year = 2025
}