Downloads · 30 days
13
7% of all-time downloads
ceofast/kumru-2b-lora
kumru-2b-lora is a text generation model from ceofast. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
This repository provides a LoRA adapter distilled from the VNGRS Kumru-2B model ( vngrs-ai/Kumru-2B, the SFT/chat variant) to be applied on top of the base model vngrs-ai/Kumru-2B-Base. The goal is to transfer Kumru’s…
Downloads · 30 days
13
7% of all-time downloads
All-time downloads
182
Public
Repo size
2.8 GB
Likes
0
Public
Click a slice to open those files.
.safetensors1.6 GB · 100%
From the Hugging Face model README
This repository provides a LoRA adapter distilled from the VNGRS Kumru-2B model (
vngrs-ai/Kumru-2B, the SFT/chat variant) to be applied on top of the base model
vngrs-ai/Kumru-2B-Base. The goal is to transfer Kumru’s chat/instruction behavior
to Kumru-2B-Base deployments with a lightweight file footprint.
vngrs-ai/Kumru-2B-Basevngrs-ai/Kumru-2B (SFT/chat)adapter_config.json + adapter_model.safetensorsfrom peft import PeftModel
from transformers import AutoModelForCausalLM, AutoTokenizer
base = "vngrs-ai/Kumru-2B-Base"
adapter = "ceofast/kumru-2b-lora"
tokenizer = AutoTokenizer.from_pretrained(base)
model = AutoModelForCausalLM.from_pretrained(base, torch_dtype="auto", device_map="auto")
model = PeftModel.from_pretrained(model, adapter, device_map="auto")
messages = [
{"role": "system", "content": "Adın Kumru, Türkçe konuşan yardımcı bir modelsin."},
{"role": "user", "content": "İstanbul'un fethi hakkında kısa bilgi verir misin?"}
]
inputs = tokenizer.apply_chat_template(messages, return_tensors="pt", add_generation_prompt=True).to(model.device)
outputs = model.generate(inputs, max_new_tokens=512, temperature=0.7, top_p=0.9)
print(tokenizer.decode(outputs[0][inputs.shape[-1]:], skip_special_tokens=True))
Not: This adapter must be used together with
vngrs-ai/Kumru-2B-Base.
The adapter is obtained by computing the delta between the base and the SFT checkpoints and factorizing it with SVD into low-rank components. In this release, the measured reconstruction error is approximately 0.409. To better preserve quality, you may increase rank/alpha and export a new version (e.g., rank 1024 / alpha 2048). A lower-error build will be added as soon as possible.