Downloads · 30 days
0
makotonlo/LLM2026_DPO_finalv5
LLM2026_DPO_finalv5 is a text generation model from makotonlo. Use it when you need the model to write or continue text. It is set up for transformers. The card lists the license as apache-2.0.
This model is a fine-tuned version of unsloth/Qwen2.5-7B-Instruct-bnb-4bit using Direct Preference Optimization (DPO) via the Unsloth library.
Downloads · 30 days
0
Access
Public
Updated Feb 8, 2026
Repo size
657 MB
Likes
0
Public
Click a slice to open those files.
.safetensors646 MB · 98%
From the Hugging Face model README
This model is a fine-tuned version of unsloth/Qwen2.5-7B-Instruct-bnb-4bit using Direct Preference Optimization (DPO) via the Unsloth library.
This repository contains LoRA adapter weights only. The base model must be loaded separately.
This repository contains LoRA adapter weights only. The base model must be loaded separately using the provided inference code.
This is a LoRA adapter model. Use it with the base model using the PEFT library.
from transformers import AutoModelForCausalLM, AutoTokenizer
import torch
model_id = f"{repo_id}" # 自動的に今回のリポジトリ名が入るように修正
tokenizer = AutoTokenizer.from_pretrained(model_id)
model = AutoModelForCausalLM.from_pretrained(
model_id,
torch_dtype=torch.float16,
device_map="auto"
)
# Test inference
prompt = "Your question here"
inputs = tokenizer.apply_chat_template([{"role": "user", "content": prompt}], tokenize=True, add_generation_prompt=True, return_tensors="pt").to("cuda")
outputs = model.generate(**inputs, max_new_tokens=512)
print(tokenizer.decode(outputs[0]))