Downloads · 30 days
11
39% of all-time downloads
jasonkang14/results
results is a machine learning model from jasonkang14. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
11
39% of all-time downloads
All-time downloads
28
Public
Repo size
168 MB
Likes
0
Public
Click a slice to open those files.
.safetensors168 MB · 95%
From the Hugging Face model README
This model is a fine-tuned version of meta-llama/Meta-Llama-3-8B-Instruct on an unknown dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Rewards/chosen | Rewards/rejected | Rewards/accuracies | Rewards/margins | Logps/rejected | Logps/chosen | Logits/rejected | Logits/chosen | Nll Loss | Log Odds Ratio | Log Odds Chosen |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 1.0958 | 0.2020 | 25 | 1.4462 | -0.1303 | -0.1689 | 0.8000 | 0.0386 | -1.6887 | -1.3025 | -1.1831 | -0.9153 | 1.3976 | -0.4854 | 0.5211 |
| 1.2563 | 0.4040 | 50 | 1.2116 | -0.1057 | -0.1387 | 0.8000 | 0.0330 | -1.3872 | -1.0575 | -1.2714 | -1.0083 | 1.1616 | -0.5001 | 0.4913 |
| 1.3121 | 0.6061 | 75 | 1.1251 | -0.0952 | -0.1249 | 0.9000 | 0.0297 | -1.2491 | -0.9524 | -1.3022 | -1.0390 | 1.0740 | -0.5109 | 0.4726 |
| 1.3689 | 0.8081 | 100 | 1.0888 | -0.0906 | -0.1196 | 0.8000 | 0.0290 | -1.1960 | -0.9064 | -1.3083 | -1.0410 | 1.0376 | -0.5127 | 0.4766 |