Downloads · 30 days
2.4K
3% of all-time downloads
syssec-utd/py37-pylingual-v1-segmenter
py37-pylingual-v1-segmenter is a token classification model from syssec-utd. Use it when you need labels on individual words, such as names. It is set up for transformers.
should probably proofread and complete it, then remove this comment. --
Downloads · 30 days
2.4K
3% of all-time downloads
All-time downloads
80.9K
Public
Parameters
108M
868 MB on disk
Likes
0
Public
Click a slice to open those files.
.safetensors434 MB · 100%
From the Hugging Face model README
This model is a fine-tuned version of syssec-utd/py37-pylingual-v1-mlm on the syssec-utd/segmentation-py37-pylingual-v1-tokenized dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Precision | Recall | F1 | Accuracy |
|---|---|---|---|---|---|---|---|
| 0.0054 | 1.0 | 94997 | 0.0036 | 0.9918 | 0.9959 | 0.9938 | 0.9983 |
| 0.0032 | 2.0 | 189994 | 0.0042 | 0.9947 | 0.9964 | 0.9956 | 0.9987 |