Downloads · 30 days
35
16% of all-time downloads
ayousanz/kiritan
kiritan is a text-to-image model from ayousanz. Use it when you need an image from a text prompt. It is set up for diffusers. The card lists the license as openrail++.
Downloads · 30 days
35
16% of all-time downloads
All-time downloads
222
Public
Repo size
383 MB
Likes
0
Public
Click a slice to open those files.
.safetensors372 MB · 97%
From the Hugging Face model README
kiritan.safetensors here 💾.
models/Lora folder.<lora:kiritan:1> to your prompt. On ComfyUI just load it as a regular LoRA.kiritan_emb.safetensors here 💾.
embeddings folderkiritan_emb to your prompt. For example, A kiritan_emb character
(you need both the LoRA and the embeddings as they were trained together for this LoRA)from diffusers import AutoPipelineForText2Image
import torch
from huggingface_hub import hf_hub_download
from safetensors.torch import load_file
pipeline = AutoPipelineForText2Image.from_pretrained('stabilityai/stable-diffusion-xl-base-1.0', torch_dtype=torch.float16).to('cuda')
pipeline.load_lora_weights('ayousanz/kiritan', weight_name='pytorch_lora_weights.safetensors')
embedding_path = hf_hub_download(repo_id='ayousanz/kiritan', filename='kiritan_emb.safetensors' repo_type="model")
state_dict = load_file(embedding_path)
pipeline.load_textual_inversion(state_dict["clip_l"], token=["<s0>", "<s1>"], text_encoder=pipeline.text_encoder, tokenizer=pipeline.tokenizer)
pipeline.load_textual_inversion(state_dict["clip_g"], token=["<s0>", "<s1>"], text_encoder=pipeline.text_encoder_2, tokenizer=pipeline.tokenizer_2)
image = pipeline('A <s0><s1> character').images[0]
For more details, including weighting, merging and fusing LoRAs, check the documentation on loading LoRAs in diffusers
To trigger image generation of trained concept(or concepts) replace each concept identifier in you prompt with the new inserted tokens:
to trigger concept TOK → use <s0><s1> in your prompt
All Files & versions.
The weights were trained using 🧨 diffusers Advanced Dreambooth Training Script.
LoRA for the text encoder was enabled. False.
Pivotal tuning was enabled: True.
Special VAE used for training: madebyollin/sdxl-vae-fp16-fix.