Downloads · 30 days
3
11% of all-time downloads
khushkaran/punjab_bert
punjab_bert is a fill-mask model from khushkaran. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as cc-by-4.0.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
3
11% of all-time downloads
All-time downloads
28
Public
Repo size
13.3 GB
Likes
0
Public
Click a slice to open those files.
.bin951 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of l3cube-pune/punjabi-bert on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 3.8977 | 1.0 | 998 | 3.7757 |
| 3.7502 | 2.0 | 1996 | 3.7404 |
| 3.7475 | 3.0 | 2994 | 3.7493 |
| 3.7423 | 4.0 | 3992 | 3.7376 |
| 3.731 | 5.0 | 4990 | 3.7338 |
| 3.7254 | 6.0 | 5988 | 3.7195 |
| 3.7213 | 7.0 | 6986 | 3.7129 |