Downloads · 30 days
12
41% of all-time downloads
prakharsinghal/lora_finetune
lora_finetune is a text generation model from prakharsinghal. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
12
41% of all-time downloads
All-time downloads
29
Public
Repo size
28.8 MB
Likes
0
Public
Click a slice to open those files.
.json28.4 MB · 58%
From the Hugging Face model README
This model is a fine-tuned version of Qwen/Qwen2.5-0.5B-Instruct on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.8734 | 0.9988 | 1000 | 0.8738 |
| 0.8169 | 1.9968 | 2000 | 0.8689 |
| 0.8325 | 2.9948 | 3000 | 0.8675 |