Downloads · 30 days
11
31% of all-time downloads
Incomple/Llama-3.1-8B-Instruct_sft_sg_values_resp_split
Llama-3.1-8B-Instruct_sft_sg_values_resp_split is a machine learning model from Incomple. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as llama3.1.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
11
31% of all-time downloads
All-time downloads
36
Public
Repo size
101 MB
Likes
0
Public
Click a slice to open those files.
.safetensors83.9 MB · 83%
From the Hugging Face model README
This model is a fine-tuned version of meta-llama/Llama-3.1-8B-Instruct on the sft_sg_values_res_split dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.4139 | 0.1710 | 250 | 0.3111 |
| 0.2133 | 0.3419 | 500 | 0.2000 |
| 0.1749 | 0.5129 | 750 | 0.1754 |
| 0.161 | 0.6839 | 1000 | 0.1615 |
| 0.1455 | 0.8548 | 1250 | 0.1563 |