Downloads · 30 days
10
20% of all-time downloads
ryomac/lora_sft_full_ep2
lora_sft_full_ep2 is a text generation model from ryomac. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
- Model type: LoRA adapter (PEFT) - Base model: Qwen/Qwen3-4B-Instruct-2507 - Goal: Structured output performance improvement for JSON-style competition prompts - Author: ryomac
Downloads · 30 days
10
20% of all-time downloads
All-time downloads
49
Public
Repo size
4.8 GB
Likes
0
Public
Click a slice to open those files.
.pt3.2 GB · 59%
From the Hugging Face model README
ryomac/lora_sft_full_ep2Qwen/Qwen3-4B-Instruct-2507u-10bei/structured_data_with_cot_dataset_512_v2full_ep2)from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel
base_id = "Qwen/Qwen3-4B-Instruct-2507"
adapter_id = "ryomac/lora_sft_full_ep2"
tokenizer = AutoTokenizer.from_pretrained(base_id)
base_model = AutoModelForCausalLM.from_pretrained(base_id)
model = PeftModel.from_pretrained(base_model, adapter_id)
For reproducibility details, refer to training/inference scripts under scripts/ in this project.