Downloads · 30 days
14
35% of all-time downloads
NasimB/bert-concat-3
bert-concat-3 is a fill-mask model from NasimB. Use it when you need the model to fill a missing word. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
14
35% of all-time downloads
All-time downloads
40
Public
Repo size
8.3 GB
Likes
0
Public
Click a slice to open those files.
.bin438 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of on the generator dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 6.5215 | 2.11 | 1000 | 6.1057 |
| 5.9958 | 4.22 | 2000 | 6.0199 |
| 5.9066 | 6.33 | 3000 | 5.9833 |
| 5.8449 | 8.44 | 4000 | 5.9594 |
| 5.7913 | 10.55 | 5000 | 5.9176 |
| 5.7418 | 12.66 | 6000 | 5.8949 |
| 5.6901 | 14.77 | 7000 | 5.8753 |
| 5.6485 | 16.88 | 8000 | 5.8592 |
| 5.6238 | 18.99 | 9000 | 5.8509 |
| 5.6704 | 21.1 | 10000 | 5.8856 |
| 5.6375 | 23.21 | 11000 | 5.8703 |
| 5.6039 | 25.32 | 12000 | 5.8635 |
| 5.5756 | 27.43 | 13000 | 5.8533 |
| 5.5437 | 29.54 | 14000 | 5.8408 |
| 5.5189 | 31.65 | 15000 | 5.8154 |
| 5.4982 | 33.76 | 16000 | 5.8028 |