Downloads · 30 days
28
65% of all-time downloads
PracticalWork/Qwen3-1.7B-tuned
Qwen3-1.7B-tuned is a text generation model from PracticalWork. Use it when you need the model to write or continue text. It is set up for peft. The card lists the license as apache-2.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
28
65% of all-time downloads
All-time downloads
43
Public
Repo size
11.6 MB
Likes
0
Public
Click a slice to open those files.
.json14.2 MB · 88%
From the Hugging Face model README
This model is a fine-tuned version of Qwen/Qwen3-1.7B on an unknown dataset. It achieves the following results on the evaluation set:
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Perplexity |
|---|---|---|---|---|
| No log | 0 | 0 | 6.3047 | 547.1235 |
| No log | 0.6011 | 333 | 1.8454 | 6.3306 |
| 1.9738 | 1.2022 | 666 | 1.7511 | 5.7610 |
| 1.9738 | 1.8032 | 999 | 1.6936 | 5.4388 |
| 1.7084 | 2.4043 | 1332 | 1.6532 | 5.2239 |
| 1.7084 | 3 | 1664 | 1.6231 | 5.0689 |