Downloads · 30 days
6
8% of all-time downloads
WDong/7B_lora_06051615
7B_lora_06051615 is a machine learning model from WDong. Use it for the machine learning task on the model card, and read the license before you ship it in a product. It is set up for peft. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
6
8% of all-time downloads
All-time downloads
75
Public
Repo size
8.4 MB
Likes
0
Public
Click a slice to open those files.
.safetensors8.4 MB · 62%
From the Hugging Face model README
This model is a fine-tuned version of Qwen/Qwen1.5-7B-Chat on the my own dataset. It achieves the following results on the evaluation set:
Qwen1.5 is the beta version of Qwen2, a transformer-based decoder-only language model pretrained on a large amount of data. In comparison with the previous released Qwen, the improvements include:
trust_remote_code.
For more details, please refer to the blog post and GitHub repo.More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.7655 | 0.4793 | 700 | 0.9256 |
| 0.8703 | 0.9586 | 1400 | 0.9017 |
| 0.725 | 1.4379 | 2100 | 0.9006 |
| 0.7958 | 1.9172 | 2800 | 0.8908 |
| 0.7346 | 2.3964 | 3500 | 0.8911 |
| 0.6516 | 2.8757 | 4200 | 0.8911 |
| 1.0524 | 3.3550 | 4900 | 0.9006 |
| 1.1005 | 3.8343 | 5600 | 0.8945 |
| 0.7991 | 4.3136 | 6300 | 0.9009 |
| 0.7668 | 4.7929 | 7000 | 0.9016 |