Downloads · 30 days
13
25% of all-time downloads
ryomac/lora_sft_v2_weak650_ep2
lora_sft_v2_weak650_ep2 is a text generation model from ryomac. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
- Model type: LoRA adapter (PEFT) for causal language modeling - Base model: Qwen/Qwen3-4B-Instruct-2507 - Purpose: Structured output tuning for JSON-style answer formatting in competition-style prompts - Developer: r…
Downloads · 30 days
13
25% of all-time downloads
All-time downloads
51
Public
Repo size
540 MB
Likes
0
Public
Click a slice to open those files.
.safetensors529 MB · 98%
From the Hugging Face model README
ryomac/lora_sft_v2_weak650_ep2Qwen/Qwen3-4B-Instruct-2507adapter_model.safetensors).u-10bei/structured_data_with_cot_dataset_512_v2docs/report_v2_weak650_ep2.md).from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel
base_id = "Qwen/Qwen3-4B-Instruct-2507"
adapter_id = "ryomac/lora_sft_v2_weak650_ep2"
tok = AutoTokenizer.from_pretrained(base_id)
base = AutoModelForCausalLM.from_pretrained(base_id)
model = PeftModel.from_pretrained(base, adapter_id)
For reproducibility context, refer to this project repository and training scripts under scripts/.