Downloads · 30 days
0
pepoo20/Orca_End
Orca_End is a machine learning model from pepoo20. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
basemodel: /home/ray/default/save/SFTmerge libraryname: peft license: other tags: - llama-factory - lora - generatedfromtrainer model-index: - name: OrcaEnd results: [] --- --
Downloads · 30 days
0
Access
Public
Updated Aug 25, 2024
Repo size
152 MB
Likes
0
Public
Click a slice to open those files.
.safetensors134 MB · 72%
From the Hugging Face model README
This model is a fine-tuned version of /home/ray/default/save/SFT_merge on the WordProblems_SFT_LLama_End dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.2267 | 0.2341 | 500 | 0.2229 |
| 0.177 | 0.4682 | 1000 | 0.1735 |
| 0.1592 | 0.7022 | 1500 | 0.1591 |
| 0.1572 | 0.9363 | 2000 | 0.1515 |
| 0.142 | 1.1704 | 2500 | 0.1472 |
| 0.1376 | 1.4045 | 3000 | 0.1441 |
| 0.1569 | 1.6386 | 3500 | 0.1425 |
| 0.1413 | 1.8727 | 4000 | 0.1417 |