Downloads · 30 days
19
29% of all-time downloads
hamini58/qwen3-4b-structeval-lora
qwen3-4b-structeval-lora is a text generation model from hamini58. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
This repository provides a LoRA adapter fine-tuned from Qwen/Qwen3-4B-Instruct-2507 for improving structured output accuracy.
Downloads · 30 days
19
29% of all-time downloads
All-time downloads
66
Public
Repo size
3.2 GB
Likes
0
Public
Click a slice to open those files.
.pt2.1 GB · 56%
From the Hugging Face model README
This repository provides a LoRA adapter fine-tuned from
Qwen/Qwen3-4B-Instruct-2507 for improving structured output accuracy.
⚠️ This repository contains LoRA adapter weights only.
The base model must be downloaded separately.
This LoRA adapter is trained to improve the model’s ability to generate strictly structured outputs, such as:
During training:
Output: marker is supervisedThis design improves format correctness without exposing or overfitting internal reasoning traces.
from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel
import torch
base_model = "Qwen/Qwen3-4B-Instruct-2507"
adapter_model = "hamini58/qwen3-4b-structeval-lora"
tokenizer = AutoTokenizer.from_pretrained(base_model)
model = AutoModelForCausalLM.from_pretrained(
base_model,
torch_dtype=torch.float16,
device_map="auto",
)
model = PeftModel.from_pretrained(model, adapter_model)
model.eval()
Training dataset:
u-10bei/structured_data_with_cot_dataset_512_v2
Dataset License:
MIT License
Compliance:
Users must comply with:
This repository distributes only LoRA adapter weights and does not redistribute the base model.