Downloads · 30 days
6
4% of all-time downloads
Harshatheeswar/Ltg_bert
Ltg_bert is a fill-mask model from Harshatheeswar. Use it when you need the model to fill a missing word. It is set up for transformers. The card lists the license as mit.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
6
4% of all-time downloads
All-time downloads
136
Public
Repo size
18 GB
Likes
0
Public
Click a slice to open those files.
.bin418 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of babylm/ltgbert-10m-2024 on the None dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss |
|---|---|---|---|
| 0.0001 | 0.9997 | 2779 | 0.0000 |
| 0.0 | 1.9997 | 5559 | 0.0000 |
| 0.0 | 2.9998 | 8339 | 0.0000 |
| 0.0 | 3.9998 | 11119 | 0.0000 |
| 0.0 | 4.9984 | 13895 | 0.0000 |