Downloads · 30 days
12
7% of all-time downloads
PracticalWork/Llama-3.2-1B-Instruct-tuned
Llama-3.2-1B-Instruct-tuned is a text generation model from PracticalWork. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as llama3.2.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
12
7% of all-time downloads
All-time downloads
180
Public
Repo size
17.4 MB
Likes
0
Public
Click a slice to open those files.
.json17.3 MB · 99%
From the Hugging Face model README
This model is a fine-tuned version of meta-llama/Llama-3.2-1B-Instruct on an unknown dataset. It achieves the following results on the evaluation set:
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Perplexity |
|---|---|---|---|---|
| No log | 0 | 0 | 4.5699 | 96.5354 |
| No log | 0.6011 | 333 | 1.7253 | 5.6141 |
| 1.8003 | 1.2022 | 666 | 1.6644 | 5.2825 |
| 1.8003 | 1.8032 | 999 | 1.6342 | 5.1252 |
| 1.6299 | 2.4043 | 1332 | 1.6094 | 5.0000 |
| 1.6299 | 3 | 1664 | 1.5964 | 4.9351 |