Downloads · 30 days
4
16% of all-time downloads
codesiddhant/lora-llama-3.2b-instruct
lora-llama-3.2b-instruct is a machine learning model from codesiddhant. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as llama3.2.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
4
16% of all-time downloads
All-time downloads
25
Public
Repo size
154 MB
Likes
0
Public
Click a slice to open those files.
.json34.5 MB · 83%
From the Hugging Face model README
This model is a fine-tuned version of meta-llama/Llama-3.2-1B-Instruct on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.119 | 0.3125 | 50 | 0.1116 |
| 0.0695 | 0.625 | 100 | 0.0709 |
| 0.049 | 0.9375 | 150 | 0.0506 |
| 0.05 | 1.25 | 200 | 0.0436 |
| 0.0367 | 1.5625 | 250 | 0.0398 |
| 0.033 | 1.875 | 300 | 0.0376 |
| 0.03 | 2.1875 | 350 | 0.0362 |
| 0.0312 | 2.5 | 400 | 0.0355 |
| 0.0299 | 2.8125 | 450 | 0.0351 |